Reward-free Pretraining for Reinforcement Learning via Occupancy Coverage Maximization
通过占据覆盖最大化进行强化学习的无奖励预训练
Marco Pratticò, Pietro Novelli, Massimiliano Pontil, Carlo Ciliberto
机构
*
Computational Statistics and Machine Learning - Istituto Italiano di Tecnologia(计算统计与机器学习 - 意大利技术研究院)
;
AI Centre, Computer Science Department, University College London(人工智能中心,计算机科学系,伦敦大学学院)
SHIELD: Safety on Humanoids via CBFs In Expectation on Learned Dynamics
SHIELD: 基于学习动力学期望的控制障碍函数实现人形机器人安全
Lizhi Yang, Blake Werner, Ryan K. Cosner, David Fridovich-Keil, Preston Culbertson, Aaron D. Ames
机构
*
Mechanical and Civil Engineering, California Institute of Technology(加州理工学院机械与土木工程系)
;
Aerospace Engineering and Engineering Mechanics, UT Austin(德克萨斯大学奥斯汀分校航空航天工程与工程力学系)
;
Computer Science, Cornell University(康奈尔大学计算机科学系)
CommentsAccepted to the 2025 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS 2025). Copyright transferred to IEEE. Video at https://youtu.be/-Qv1wR4jfj4
机构
*
State Key Laboratory of Multimedia Information Processing, School of Computer Science, Peking University(北京大学计算机学院多媒体信息处理国家重点实验室)
;
Northeastern University(东北大学)
;
Tsinghua University(清华大学)
Real-Time Reinforcement Learning for Dynamic Tasks with a Parallel Soft Robot
动态任务的实时强化学习与并行软机器人
James Avtges, Jake Ketchum, Millicent Schlafly, Helena Young, Taekyoung Kim, Allison Pinosky, Ryan L. Truby, Todd D. Murphey
机构
*
Department of Mechanical Engineering, Northwestern University(西北大学机械工程系)
;
Department of Materials Science and Engineering, Northwestern University(西北大学材料科学与工程系)
Coherent Off-Policy Improvement of Large Behavior Models with Learned Rewards
基于学习奖励的大规模行为模型的一致性离策略改进
Christian Scherer, Joe Watson, Theo Gruner, Daniel Palenicek, Ingmar Posner, Jan Peters
机构
*
Technical University of Darmstadt(达姆施塔特技术大学)
;
University of Oxford(牛津大学)
;
Zuse School ELIZA(泽努斯学校ELIZA)
;
hessian.AI(海西斯AI)
;
German Research Center for AI (DFKI)(德国人工智能研究中心(DFKI))
;
Robotics Institute Germany (RIG)(德国机器人研究所)
机构
*
College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院)
;
State Key Laboratory of Transvascular Implantation Devices of the Second Affiliated Hospital, Zhejiang University School of Medicine(浙江大学医学院第二附属医院血管植入设备国家重点实验室)
;
Dessight Biomedical(Dessight生物医学公司)
;
Center for Rehabilitation Medicine, Department of Ophthalmology, Zhejiang Provincial People’s Hospital(浙江省人民医院康复医学中心、眼科部门)
;
School of Biosystems Engineering and Food Science, Zhejiang University(浙江大学生物系统工程与食品科学学院)
;
School of Public Health and Second Affiliated Hospital, Zhejiang University School of Medicine(浙江大学医学院公共卫生学院及第二附属医院)
;
State Key Laboratory of Transvascular Implantation Devices of the Second Affiliated Hospital and School of Public Health, Zhejiang University School of Medicine(浙江大学医学院第二附属医院及公共卫生学院血管植入设备国家重点实验室)
;
Zhejiang Key Laboratory of Medical Imaging Artificial Intelligence(浙江省医学影像人工智能重点实验室)
Domain-Adaptable Reinforcement Learning for Code Generation with Dense Rewards
用于密集奖励的领域可适应强化学习代码生成
Erfan Aghadavoodi Jolfaei, Daniel Maninger, Abhinav Anand, Mert Tiftikci, Mira Mezini
机构
*
Hessian Center for Artificial Intelligence (hessian.AI)(海斯曼人工智能中心)
;
National Research Center for Applied Cybersecurity ATHENE(应用网络安全国家研究中心ATHENE)
Test-time Offline Reinforcement Learning on Goal-related Experience
测试时的离线强化学习在目标相关经验上的应用
Marco Bagatella, Mert Albaba, Jonas Hübotter, Georg Martius, Andreas Krause
机构
*
ETH Zurich, Zurich, Switzerland(苏黎世联邦理工学院,苏黎世,瑞士)
;
Max Planck Institute for Intelligent Systems, Tubingen, Germany(智能系统马克斯·普朗克研究所,图宾根,德国)
;
University of Tubingen, Tubingen, Germany(图宾根大学,图宾根,德国)
机构
*
School of Computation Information and Technology, Technical University of Munich(技术大学慕尼黑计算信息与技术学院)
;
Department of Computer Science, University of Innsbruck(因斯布鲁克大学计算机科学系)
;
L3S Research Center, Leibniz Universität Hannover(汉诺威莱布尼茨大学L3S研究中心)
;
Digital Science Center (DiSC), University of Innsbruck(因斯布鲁克大学数字科学中心)
机构
*
School of Computer Science and Engineering, Sun Yat-sen University(中山大学计算机科学与工程学院)
;
Guangdong Key Laboratory of Big Data Analysis and Processing(广东省大数据分析与处理重点实验室)
机构
*
University of Missouri–Kansas City(密苏里大学堪萨斯城分校)
;
Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
U. S. Naval Research Laboratory(美国海军研究实验室)
;
Lamar University(拉马尔大学)
;
Meta AI
;
Rochester Institute of Technology(罗彻斯特理工学院)