Learning Robust Execution in Robotic Manipulation with Agentic Reinforcement Learning
通过智能强化学习在机器人操作中学习鲁棒执行
Xiaopeng Zhang, Yueyang Weng, Qi Liu, Yongjin Mu, Yanjie Li
机构
*
School of Inteligence Science and Engineering, the Harbin Institute of Technology Shenzhen(哈尔滨工业大学(深圳)智能科学与工程学院)
;
Faculty of Robot Science and Engineering, Northeastern University(东北大学机器人科学与工程学院)
Learning More from Less: Reinforcement Learning from Hindsight
从更少中学习更多:事后诸葛亮式强化学习
Iris Xu, Sunshine Jiang, John Marangola, Nitish Dashora, Richard Li, Thomas Liu, Zexue He, Yuheng Zhi, Alex Pentland, Pulkit Agrawal, Zhang-Wei Hong
机构
*
Massachusetts Institute of Technology(麻省理工学院)
;
MIT-IBM Computing Research Lab(麻省理工学院-IBM计算研究实验室)
;
Stanford University(斯坦福大学)
;
University of California, San Diego(加利福尼亚大学圣地亚哥分校)
Human-as-Humanoid: Enabling Zero-Shot Humanoid Learning from Ego-Exo Human Videos with Human-Aligned Embodiments
人类作为人形机器人:通过人类对齐的具身从自我-外部人类视频实现零样本人形机器人学习
Xiaopeng Lin, Ruoqi Yang, Shijie Lian, Zhaolong Shen, Bin Yu, Changti Wu, Haibao Liu, Yuxiang Zhang, Hong Li, Qiyuan Su, Haochen Liu, Xuguo He, Yukun Shi, Cong Huang, Zhirui Zhang, Bojun Cheng, Kai Chen
机构
*
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
DeepCybo
;
ZGCA
;
ZGCI
;
Harbin Institute of Technology(哈尔滨工业大学)
;
Huazhong University of Science and Technology(华中科技大学)
;
Beihang University(北京航空航天大学)
ReactVLA: Fast and Lightweight Reactive Robot Manipulation via Improved Mean Flow Action Generation
ReactVLA: 通过改进的平均流动作生成实现快速轻量级反应式机器人操作
Yanzhao Guo, Wenkai Chen, Jianwei Zhang
机构
*
Shanghai Jiao Tong University(上海交通大学)
;
Technical Aspects of Multimodal Systems (TAMS), Department of Informatics, Universität Hamburg(汉堡大学信息学系多模态系统技术方面(TAMS))
机构
*
Shanghai Jiao Tong University(上海交通大学)
;
Alibaba Group(阿里巴巴集团)
;
Tianji KernalMind Co., Ltd.(天机芯智有限公司)
;
The Hong Kong University of Science and Technology(香港科技大学)
;
Southeast University(东南大学)
;
Renmin University of China(中国人民大学)
;
The University of Tokyo(东京大学)
QPILOTS: Efficient Test-Time Q-Steering for Flow Policies
QPILOTS:面向流策略的高效测试时Q引导
Yifan Ruan, Chenyang Cao, Andreas Burger, Ali Pesaranghader, Kaveh Kamali, Jaehong Kim, Nandita Vijaykumar, Alan Aspuru-Guzik, Igor Gilitschenski, Nicholas Rhinehart
机构
*
University of Toronto(多伦多大学)
;
Vector Institute(向量研究所)
;
LG Electronics(LG电子)
E-TTS: A New Embodied Test-Time Scaling Framework for Robotic Manipulation
E-TTS:一种新的机器人操作具身测试时缩放框架
Wen Ye, Peiyan Li, Tingyu Yuan, Yuan Xu, Xiangnan Wu, Chaoyang Zhao, Jing Liu, Nianfeng Liu, Yan Huang, Liang Wang
机构
*
New Laboratory of Pattern Recognition (NLPR), Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所模式识别国家重点实验室)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
Foundation Model Research Center, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所基础模型研究中心)
;
FiveAges(五时代)