机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Chongqing Chang’an Technology Co., Ltd(重庆长安科技有限公司)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
College of AI, Tsinghua University(清华大学人工智能学院)
;
Zhongguancun Academy(中关村学院)
Utilizing Vision-Language Models as Action Models for Intent Recognition and Assistance
Cesar Alan Contreras, Manolis Chiou, Alireza Rastegarpanah, Michal Szulik, Rustam Stolkin
机构
*
University of Birmingham(伯明翰大学)
;
Queen Mary University of London(伦敦大学女王学院)
;
Aston University(阿斯顿大学)
;
United Kingdom National Nuclear Laboratory Ltd.(英国国家核实验室有限公司)
专题命中
VLA模型
:action model(title);分类 cs.RO、cs.AI
CommentsAccepted at Human-Centered Robot Autonomy for Human-Robot Teams (HuRoboT) at IEEE RO-MAN 2025, Eindhoven, the Netherlands
QPILOTS: Efficient Test-Time Q-Steering for Flow Policies
QPILOTS:面向流策略的高效测试时Q引导
Yifan Ruan, Chenyang Cao, Andreas Burger, Ali Pesaranghader, Kaveh Kamali, Jaehong Kim, Nandita Vijaykumar, Alan Aspuru-Guzik, Igor Gilitschenski, Nicholas Rhinehart
机构
*
University of Toronto(多伦多大学)
;
Vector Institute(向量研究所)
;
LG Electronics(LG电子)
Explainable deep learning improves human mental models of self-driving cars
可解释深度学习提升人类对自动驾驶汽车的心理模型
Eoin M. Kenny, Akshay Dharmavaram, Sang Uk Lee, Tung Phan-Minh, Shreyas Rajesh, Yunqing Hu, Laura Major, Momchil S. Tomov, Julie A. Shah
机构
*
Computer Science & Artificial Intelligence Laboratory (CSAIL), Massachusetts Institute of Technology(计算机科学与人工智能实验室(CSAIL),麻省理工学院)
;
Motional AD Inc.(Motional AD公司)
;
Department of Psychology and Center for Brain Science, Harvard University(心理学系和大脑科学中心,哈佛大学)
;
Department of Aeronautics and Astronautics, Massachusetts Institute of Technology(航空与宇航系,麻省理工学院)
Supervised Mixture-of-Experts for Surgical Grasping and Retraction
监督混合专家架构用于手术抓取与牵开
Lorenzo Mazza, Ariel Rodriguez, Rayan Younis, Martin Lelis, Ortrun Hellig, Chenpan Li, Sebastian Bodenstedt, Martin Wagner, Stefanie Speidel
机构
*
Department of Translational Surgical Oncology, NCT/UCC Dresden, Faculty of Medicine and University Hospital Carl Gustav Carus, TUD, Dresden, Germany(translational Surgical Oncology部门,NCT/UCC Dresden,医学院和Carl Gustav Carus大学医院,TUD,德累斯顿,德国)
;
National Center for Tumor Diseases (NCT), NCT/UCC Dresden, a partnership between DKFZ, Faculty of Medicine and University Hospital Carl Gustav Carus, TUD, and HZDR, Dresden, Germany(肿瘤疾病国家中心(NCT),NCT/UCC德累斯顿,DKFZ、医学院和Carl Gustav Carus大学医院、TUD以及HZDR之间的合作,德累斯顿,德国)
;
Faculty of Computer Science, TUD, Dresden, Germany(计算机科学系,TUD,德累斯顿,德国)
;
The Center for Tactile Internet with Human-in-the-Loop (CeTI), TUD, Dresden, Germany(带有Human-in-the-Loop的触觉互联网中心(CeTI),TUD,德累斯顿,德国)
;
Department of Visceral, Thoracic and Vascular Surgery, Faculty of Medicine and University Hospital Carl Gustav Carus, TUD, Dresden, Germany(visceral、胸腔和血管外科部门,医学院和Carl Gustav Carus大学医院,TUD,德累斯顿,德国)
专题命中
VLA模型
:vision language action(abstract);action model(abstract);分类 cs.RO、cs.AI、cs.LG
Towards the Vision-Sound-Language-Action Paradigm: The HEAR Framework for Sound-Centric Manipulation
迈向视觉-声音-语言-动作范式:用于以声音为中心的操作的HEAR框架
Chang Nie, Tianchen Deng, Guangming Wang, Zhe Liu, Hesheng Wang
机构
*
School of Automation and Intelligent Sensing, Shanghai Jiao Tong University and Shanghai Key Laboratory of Navigation and Location Based Services(自动化与智能感知学院,上海交通大学,导航与基于位置的服务重点实验室)
;
Department of Engineering, Cambridge University(工程系,剑桥大学)