SAMoE-VLA: A Scene Adaptive Mixture-of-Experts Vision-Language-Action Model for Autonomous Driving
SAMoE-VLA:一种面向自动驾驶的场景自适应混合专家视觉-语言-动作模型
机构 * Institute for AI Industry Research (AIR), Tsinghua University(人工智能产业研究院(AIR),清华大学) ; School of Instrument Science and Engineering, Southeast University(仪器科学与工程学院,东南大学) ; Zhili College, Tsinghua University(紫荆学院,清华大学) ; School of Artificial Intelligence and Automation, Huazhong University of Science and Technology(人工智能与自动化学院,华中科技大学) ; Department of Automation, University of Science and Technology of China(自动化学院,中国科学技术大学) ; Department of Automation, University of Science and Technology Beijing(自动化学院,北京科技大学)
AI总结 SAMoE-VLA通过场景自适应混合专家机制提升自动驾驶中的视觉-语言-动作推理性能。