机构
*
Guangdong Laboratory of Artificial Intelligence and Digital Economy (SZ)(广东人工智能与数字经济实验室)
;
Tsinghua University(清华大学)
;
Shenzhen University(深圳大学)
;
Meituan(美团)
;
Division of AMC and Department of ECE, HKUST(HKUST AMC 分部和电子工程系)
AnchorVLA: Bridging Discrete Decisions and Continuous Trajectories for Vision-Language-Action Planning
AnchorVLA:为视觉-语言-动作规划连接离散决策与连续轨迹
Qi Liu, Yabei Li, Hongsong Wang, Heng Zhang, Lei He
机构
*
School of Vehicle and Mobility, Tsinghua University(清华大学车辆与运载学院)
;
State Key Laboratory of Intelligent Green Vehicle and Mobility, Tsinghua University(清华大学智能绿色车辆与交通国家重点实验室)
;
School of Computer Science and Engineering, Southeast University(东南大学计算机科学与工程学院)
;
Meituan Inc.(美团公司)
;
Dongfeng Motor Corporation(东风汽车公司)
机构
*
The Hong Kong Polytechnic University(香港理工大学)
;
Meituan Longcat Team(美团龙猫团队)
;
Peking University(北京大学)
;
Nanjing University of Science and Technology(南京理工大学)
机构
*
University of Chicago(芝加哥大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Stanford University(斯坦福大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Meituan(美团)
Beyond Trajectory Imitation: Strategy-Guided Policy Optimization for LLM Reasoning
超越轨迹模仿:策略引导的策略优化用于大语言模型推理
Tianyuan Shi, Canbin Huang, Bei Li, Xin Chen, Xiaojun Quan, Jingang Wang, Qifan Wang
机构
*
School of Computer Science and Engineering, Sun Yat-sen University, China(中山大学计算机科学与工程学院)
;
Meituan, Inc., China(美团)
;
Shenzhen Loop Area Institute, China(深圳环区研究所)
;
Meta AI, USA(Meta AI)
机构
*
Meituan(美团)
;
The University of Hong Kong(香港大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Nanjing University(南京大学)
;
Harbin Institute of Technology(哈尔滨工业大学)
;
Australian Institute for Machine Learning, Adelaide University(阿德莱德大学澳大利亚机器学习研究所)
;
Ludwig Maximilian University of Munich(慕尼黑大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Queen Mary University of London(伦敦玛丽女王大学)
TORL-VLA: Tactile Guided Online Reinforcement Learning for Contact-Rich Manipulation
TORL-VLA:触觉引导的在线强化学习用于接触丰富操作
Huaihang Zheng, Yi Yang, Kai Ma, Shenglin Xu, Tian Xie, Guozheng Li, Xiangyu Wang, Yiren Ma, Si Liu, Yinian Mao, Baoxu Liu
机构
*
Meituan(美团)
;
Beijing Institute of Technology(北京理工大学)
;
Beihang University(北京航空航天大学)
;
State Key Lab of Multimodal Artificial Intelligence Systems, Institute of Automation, CAS(中国科学院自动化研究所多模态人工智能系统国家重点实验室)
;
China University of Mining and Technology (Beijing)(中国矿业大学(北京))
机构
*
Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学高瓴人工智能学院)
;
Department of Data Science, City University of Hong Kong(香港城市大学数据科学系)
;
Meituan(美团)
;
WeChat, Tencent(腾讯微信)
;
Beijing Key Laboratory of Research on Large Models and Intelligent Governance(大型模型与智能治理北京市重点实验室)