COP-Q: Safety-First Reinforcement Learning for Robot Control via Cholesky-Ordered Projection
COP-Q:基于Cholesky有序投影的安全优先强化学习机器人控制
Guopeng Li, Moritz A. Zanger, Matthijs T. J. Spaan, Julian F. P. Kooij
机构
*
Department of Cognitive Robotics, Delft University of Technology(代尔夫特理工大学认知机器人系)
;
Department of Intelligent Systems, Delft University of Technology(代尔夫特理工大学智能系统系)
;
School of Transportation, Southeast University(东南大学交通学院)
Learning Stabilizable Dynamical Systems via Control Contraction Metrics
通过控制收缩度量学习可稳定化的动态系统
Sumeet Singh, Vikas Sindhwani, Jean-Jacques E. Slotine, Marco Pavone
机构
*
Dept. of Aeronautics and Astronautics, Stanford University(航空航天系,斯坦福大学)
;
Google Brain Robotics, New York(谷歌大脑机器人,纽约)
;
Dept. of Mechanical Engineering, Massachusetts Institute of Technology(机械工程系,麻省理工学院)
CommentsTo appear at WAFR 2018. v2: re-structured Sections 3 & 4 to improve clarity; expanded discussion on limitations & future work in Section 5; added details on training & validation, significantly expanded experiments
Mind Dreamer: Untethering Imagination via Active Causal Intervention on Latent Manifolds
Mind Dreamer: 通过潜在流形上的主动因果干预释放想象力
Shaojun Xu, Xiaoling Zhou, Yihan Lin, Yapeng Meng, Xinglong Ji, Luping Shi, Rong Zhao
机构
*
Center for Brain-Inspired Computing Research, Department of Precision Instrument, Tsinghua University, Beijing, China(脑启发计算研究中心,精密仪器系,清华大学,北京,中国)
;
College of Computer Science and Technology, Zhejiang University, Hangzhou, China(计算机科学与技术学院,浙江大学,杭州,中国)
;
Pen-Tung Sah Institute of Micro-Nano Science and Technology, Xiamen University, Xiamen, China(彭途萨微纳米科学与技术研究院,厦门大学,厦门,中国)
LASER: Learning Active Sensing for Continuum Field Reconstruction
LASER: 用于连续场重建的学习主动感知
Huayu Deng, Jinghui Zhong, Xiangming Zhu, Yunbo Wang, Xiaokang Yang
机构
*
MoE Key Lab of Artificial Intelligence, AI Institute, School of Computer Science, Shanghai Jiao Tong University(人工智能MOE重点实验室、人工智能研究院、计算机科学学院、上海交通大学)
Inversely Learning Transferable Rewards via Abstracted States
通过抽象状态逆向学习可迁移奖励
Yikang Gui, Prashant Doshi
机构
*
THINC Lab, School of Computing University of Georgia(THINC实验室,计算学院,佐治亚大学)
;
School of Computing and Institute for AI University of Georgia(计算学院和人工智能研究所,佐治亚大学)
Yu Luo, Shuo Han, Yihan Hu, Lei Lv, Huaping Liu, Fuchun Sun, Jianye Hao, Dong Li
机构
*
Department of Foundation Model, 2012 Labs, Huawei(华为基础模型部门,2012实验室)
;
Shanghai Research Institute for Intelligent Autonomous Systems, Tongji University(上海智能自主系统研究院,同济大学)
;
Department of Computer Science and Technology, Tsinghua University(清华大学计算机科学与技术系)
;
College of Intelligence and Computing, Tianjin University(天津大学智能与计算学院)
CommentsAccepted and published in the Proceedings of the 29th European Conference on Applications of Evolutionary Computation (EvoApplications 2026), held as part of EvoStar 2026, Toulouse, France, April 8 to 10, 2026. Lecture Notes in Computer Science (LNCS), Springer Nature Switzerland
SkillTree: Explainable Skill-Based Deep Reinforcement Learning for Long-Horizon Control Tasks
SkillTree: 面向长时域控制任务的可解释基于技能的深度强化学习
Yongyan Wen, Siyuan Li, Rongchang Zuo, Lei Yuan, Hangyu Mao, Peng Liu
机构
*
Faculty of Computing, Harbin Institute of Technology(哈尔滨工业大学计算机学院)
;
National Key Laboratory of Novel Software Technology, Nanjing University(南京大学新型软件技术国家实验室)
;
School of Artificial Intelligence, Nanjing University(南京大学人工智能学院)
;
Polixir Technologies
;
SenseTime Research(商汤科技研究院)
SENIOR: Efficient Query Selection and Preference-Guided Exploration in Preference-based Reinforcement Learning
SENIOR: 在基于偏好的强化学习中高效查询选择与偏好引导探索
Hexian Ni, Tao Lu, Haoyuan Hu, Yinghao Cai, Shuo Wang
机构
*
State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,自动化研究所)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
机构
*
School of Computer Science, Wuhan University(武汉大学计算机学院)
;
School of Integrated Circuits, Peking University(北京大学集成电路学院)
;
School of Information, Huazhong Agricultural University(华中农业大学信息学院)
;
Cyberspace Institute of Advanced Technology, Guangzhou University(广州大学先进技术网络研究院)
机构
*
The Hong Kong University of Science and Technology (GZ)(香港科技大学(广州))
;
National University of Singapore(新加坡国立大学)
;
ShanghaiTech University(上海科技大学)
;
East China Normal University(华东师范大学)
;
Nanjing University of Information Science & Technology(南京信息工程大学)
;
Zhejiang University(浙江大学)
;
Institute of Automation, Chinese Academy of Science(中国科学院自动化研究所)
;
Shanghai AI Laboratory(上海人工智能实验室)
机构
*
Robotics and AI group, in the Department of Computer Science, Electrical and Space Engineering at Luleå University of Technology, Sweden(鲁尔坎大学技术学院机器人与人工智能小组,计算机科学、电气与空间工程系,瑞典)