Behavior Cloning of MPC for 3-DOF Robotic Manipulators
三自由度机械臂MPC的行为克隆
Theo Guegan, Dexter Wen Jie Teo
机构
*
University of Waterloo(多伦多大学)
;
Universite de Technologie de Compiègne(技术与科学大学)
;
Nanyang Technological University(南洋理工大学)
;
Polytechnique Montréal(蒙特利尔理工学院)
Coherent Off-Policy Improvement of Large Behavior Models with Learned Rewards
基于学习奖励的大规模行为模型的一致性离策略改进
Christian Scherer, Joe Watson, Theo Gruner, Daniel Palenicek, Ingmar Posner, Jan Peters
机构
*
Technical University of Darmstadt(达姆施塔特技术大学)
;
University of Oxford(牛津大学)
;
Zuse School ELIZA(泽努斯学校ELIZA)
;
hessian.AI(海西斯AI)
;
German Research Center for AI (DFKI)(德国人工智能研究中心(DFKI))
;
Robotics Institute Germany (RIG)(德国机器人研究所)
机构
*
College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院)
;
State Key Laboratory of Transvascular Implantation Devices of the Second Affiliated Hospital, Zhejiang University School of Medicine(浙江大学医学院第二附属医院血管植入设备国家重点实验室)
;
Dessight Biomedical(Dessight生物医学公司)
;
Center for Rehabilitation Medicine, Department of Ophthalmology, Zhejiang Provincial People’s Hospital(浙江省人民医院康复医学中心、眼科部门)
;
School of Biosystems Engineering and Food Science, Zhejiang University(浙江大学生物系统工程与食品科学学院)
;
School of Public Health and Second Affiliated Hospital, Zhejiang University School of Medicine(浙江大学医学院公共卫生学院及第二附属医院)
;
State Key Laboratory of Transvascular Implantation Devices of the Second Affiliated Hospital and School of Public Health, Zhejiang University School of Medicine(浙江大学医学院第二附属医院及公共卫生学院血管植入设备国家重点实验室)
;
Zhejiang Key Laboratory of Medical Imaging Artificial Intelligence(浙江省医学影像人工智能重点实验室)
Task-Induced Representational Invariances Depend on Learning Objective in Deep RL
任务诱导的表征不变性依赖于深度强化学习中的学习目标
Manu Srinath Halvagal, Sebastian Lee, SueYeon Chung
机构
*
Department of Physics, Harvard University(哈佛大学物理系)
;
Kempner Institute, Harvard University(哈佛大学凯普纳研究所)
;
Center for Computational Neuroscience, Flatiron Institute(Flatiron研究所计算神经科学中心)
All Models are Wrong, Knowing Where is Useful: On Model Uncertainty in Reinforcement Learning
所有模型都是错的,知道哪里有用:强化学习中的模型不确定性
Bernd Frauenknecht, Devdutt Subhasish, Artur Eisele, Friedrich Solowjow, Sebastian Trimpe
机构
*
German Federal Ministry of Research, Technology and Space (BMFTR)(德国联邦研究、技术和空间部)
;
Robotics Institute Germany (RIG)(德国机器人研究所)
;
Institute for Data Science in Mechanical Engineering, RWTH Aachen University(机械工程数据科学研究所,亚琛工业大学)
;
NHR Center NHR4CES at RWTH Aachen University(亚琛工业大学NHR4CES中心)
Topology-Aware State Abstraction with Tangle Cores for Markov Decision Processes
基于纠缠核的马尔可夫决策过程拓扑感知状态抽象
Ibne Farabi Shihab, Sanjeda Akter, Anuj Sharma
机构
*
Department of Computer Science, Iowa State University(计算机科学系,爱荷华州立大学)
;
Department of Civil, Construction & Environmental Engineering, Iowa State University(土木、建设与环境工程系,爱荷华州立大学)