机构
*
The Chinese University of Hong Kong(香港中文大学)
;
Nagoya University(名古屋大学)
;
Institute of Science Tokyo(东京科学大学)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
机构
*
The Chinese University of Hong Kong(香港中文大学)
;
Wuhan University of Technology(武汉理工大学)
;
The Hong Kong University of Science and Technology(香港科技大学)
;
Huawei Noah’s Ark Lab(华为诺亚方舟实验室)
机构
*
National and Local Joint Engineering Laboratory of Computer Aided Design, School of Software Engineering, Dalian University(大连大学软件工程学院计算机辅助设计国家地方联合工程实验室)
;
Department of Radiology, Xinhua Hospital Affiliated to Dalian University(大连大学附属新华医院放射科)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
University of California, San Francisco(加州大学旧金山分校)
;
Yale University(耶鲁大学)
;
The Hong Kong Polytechnic University(香港理工大学)
VLAFlow: A Unified Training Framework for Vision-Language-Action Models via Co-training and Future Latent Alignment
VLAFlow:通过协同训练和未来潜在对齐的视觉-语言-动作模型统一训练框架
Guoyang Xia, Fengfa Li, Hongjin Ji, Lei Ren, Fangxiang Feng, Kun Zhan, Yan Xie
机构
*
Li Auto Inc.(理想汽车)
;
School of Artificial Intelligence, Beijing University of Posts and Telecommunications(北京邮电大学人工智能学院)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
机构
*
Hong Kong University of Science and Technology(香港科学与技术大学)
;
Zhejiang University(浙江大学)
;
National University of Singapore(新加坡国立大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Peking University(北京大学)
;
Independent Researcher(独立研究者)