机构
*
Department of Artificial Intelligence, School of Informatics, Xiamen University(厦门大学信息学院人工智能系)
;
Department of Computer Science, Aberystwyth University(阿伯里斯特威斯大学计算机科学系)
Self-Improving VLA Policies: Selected Diffusion Noise for Spurious-Robust Action Smoothing
自我改进的VLA策略:用于抗伪影动作平滑的选择性扩散噪声
Duc Minh Nguyen, Bao-Ngoc Dao, Tung M. Luu, Binh Gia Nguyen, Vinh Tong, Anji Liu, Vu N. Duong, Dung D. Le, Daniel Sonntag, Trung Le, Ngan Le, Jan Peter, An Thai Le, Minh Nhat Vu, Mathias Niepert, Khoa D. Doan, Duy M. H. Nguyen, Vien Anh Ngo
机构
*
Center for AI Research, VinUniversity(VinUniversity人工智能研究中心)
;
VinRobotics
;
KAIST(韩国科学技术院)
;
University of Stuttgart(斯图加特大学)
;
IMPRS-IS(国际马克斯·普朗克智能系统研究学院)
;
National University of Singapore(新加坡国立大学)
;
DFKI(德国人工智能研究中心)
;
University of Oldenburg(奥尔登堡大学)
;
Monash University(莫纳什大学)
;
University of Arkansas(阿肯色大学)
;
TU Darmstadt(达姆施塔特工业大学)
Do Vision-Language-Action Models Mean What They Say? On the Role of Faithfulness in Embodied Reasoning
视觉语言动作模型所言即所指?论忠实性在具身推理中的作用
Matthew Foutter, Matteo Cercola, Lena Wild, Yunshan Wang, Michelle Li, Daniele Gammelli, Marco Pavone
机构
*
Stanford University(斯坦福大学)
;
Politecnico di Milano(米兰理工大学)
;
KTH Royal Institute of Technology(皇家理工学院)
;
Italian Institute of Artificial Intelligence (AI4I)(意大利人工智能研究所)
;
NVIDIA Research(英伟达研究院)
Towards Backdoor-Based Ownership Verification for Vision-Language-Action Models
面向视觉-语言-动作模型的后门基于所有权验证
Ming Sun, Rui Wang, Xingrui Yu, Lihua Jing, Hangyu Du, Zhenglin Wan, Xu Pan, Ivor Tsang
机构
*
Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)
;
School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络安全学院)
;
A*STAR Institute of High Performance Computing (A*STAR IHPC)(新加坡A*STAR高性能计算研究所)
;
A*STAR Centre for Frontier AI Research (A*STAR CFAR)(新加坡A*STAR前沿人工智能研究中心)
;
College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算机与数据科学学院)
;
College of Design and Engineering, Nanyang Technological University(南洋理工大学设计与工程学院)
;
Department of Computer Science, National University of Singapore(新加坡国立大学计算机科学系)
;
State Key Laboratory of Information Engineering in Surveying, Mapping and Remote Sensing (LIESMARS), Wuhan University(武汉大学测绘遥感信息工程国家重点实验室(LIESMARS))
ActDistill: General Action-Guided Self-Derived Distillation for Efficient Vision-Language-Action Models
ActDistill: 通用动作引导自衍生蒸馏用于高效视觉-语言-动作模型
Wencheng Ye, Tianshi Wang, Lei Zhu, Fengling Li, Guoli Yang, Hengtao Shen
机构
*
School of Computer Science and Technology, Tongji University(同济大学计算机科学与技术学院)
;
University of Technology Sydney(悉尼科技大学)
;
Advanced Institute of Big Data(大数据先进研究院)
机构
*
Thomas Lord Department of Computer Science, University of Southern California(南加州大学托马斯·洛德计算机科学系)
;
Sony AI(索尼AI)
;
Sibley School of Mechanical and Aerospace Engineering, Cornell University(康奈尔大学西布利机械与航空航天工程学院)
机构
*
East China Normal University(华东师范大学)
;
Zhongguancun Academy(中关村学院)
;
CFAR, A*STAR, Singapore(新加坡科技研究局CFAR)
;
Tsinghua University(清华大学)
;
Harbin Institute of Technology(哈尔滨工业大学)