A Pontryagin Method of Model-based Reinforcement Learning via Hamiltonian Actor-Critic
基于汉密尔顿量的模型驱动强化学习方法:通过汉密尔顿演员-评论家方法
Chengyang Gu, Yuxin Pan, Hui Xiong, Yize Chen
机构
*
Information Hub, HKUST (Guangzhou)(香港科技大学(广州)信息枢纽)
;
Department of Computer Science, City University of Hong Kong(香港城市大学计算机科学系)
;
Department of Electrical and Computer Engineering, University of Alberta(阿尔伯塔大学电气与计算机工程系)
Neural ODE and SDE Models for Adaptation and Planning in Model-Based Reinforcement Learning
神经ODE和SDE模型用于基于模型的强化学习中的适应与规划
Chao Han, Stefanos Ioannou, Luca Manneschi, T. J. Hayward, Michael Mangan, Aditya Gilra, Eleni Vasilaki
机构
*
School of Computer Science, The University of Sheffield, UK(谢菲尔德大学计算机科学学院,英国)
;
Cancer Research UK National Biomarker Centre, The University of Manchester, UK(英国曼彻斯特大学癌症研究UK国家生物标志物中心)
;
School of Chemical, Materials and Biological Engineering, The University of Sheffield, UK(谢菲尔德大学化学、材料和生物工程学院,英国)
;
Machine Learning group, Centrum Wiskunde & Informatica, Amsterdam, Netherlands(荷兰阿姆斯特丹Centrum Wiskunde & Informatica机器学习组)
;
Institute for Ecological Economics, Vienna University of Economics and Business, Austria(奥地利维也纳经济与商业大学生态经济研究所)
Model-Based Reinforcement Learning Under Confounding
基于模型的强化学习中的混杂问题
Nishanth Venkatesh, Andreas A. Malikopoulos
机构
*
Department of Systems Engineering, Cornell University(系统工程系,康奈尔大学)
;
School of Civil and Environmental Engineering, Cornell University(土木与环境工程学院,康奈尔大学)