SAMoE-VLA: A Scene Adaptive Mixture-of-Experts Vision-Language-Action Model for Autonomous Driving
SAMoE-VLA:一种面向自动驾驶的场景自适应混合专家视觉-语言-动作模型
Zihan You, Hongwei Liu, Chenxu Dang, Zhe Wang, Sining Ang, Aoqi Wang, Yan Wang
机构
*
Institute for AI Industry Research (AIR), Tsinghua University(人工智能产业研究院(AIR),清华大学)
;
School of Instrument Science and Engineering, Southeast University(仪器科学与工程学院,东南大学)
;
Zhili College, Tsinghua University(紫荆学院,清华大学)
;
School of Artificial Intelligence and Automation, Huazhong University of Science and Technology(人工智能与自动化学院,华中科技大学)
;
Department of Automation, University of Science and Technology of China(自动化学院,中国科学技术大学)
;
Department of Automation, University of Science and Technology Beijing(自动化学院,北京科技大学)
机构
*
Tuojing Intelligence
;
The University of Hong Kong(香港大学)
;
King's College London(伦敦国王学院)
;
The University of Sydney(悉尼大学)
;
Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)