Directional Constraints for Efficient Exploration in Safe Reinforcement Learning
安全强化学习中高效探索的方向约束
Paolo Magliano, Puze Liu, Jan Peters, Davide Tateo, Raffaello Camoriano
机构
*
Dipartimento di Automatica e Informatica, Politecnico di Torino(自动控制与信息工程系,都灵理工大学)
;
Tongji University, Shanghai Research Institute for Intelligent Autonomous Systems(同济大学,上海智能自主系统研究院)
;
German Research Center for AI (DFKI)(德国人工智能研究中心(DFKI))
;
Intelligent Autonomous Systems Group, TU Darmstadt(智能自主系统组,达姆施塔特工业大学)
;
Lund University(隆德大学)
;
Istituto Italiano di Tecnologia(意大利技术研究院)
CommentsThis paper has been accepted for publication at the IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), Pittsburgh, USA, 2026. 8 pages, 8 figures
机构
*
Monash University(莫纳什大学)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Shenzhen University of Advanced Technology(深圳先进技术大学)
;
Pusan National University(釜山国立大学)
机构
*
South China University of Technology(华南理工大学)
;
Westlake University(西湖大学)
;
Johns Hopkins University(约翰·霍普金斯大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Shenzhen Loop Area Institute(深圳河套学院)