Comments16 pages, 9 figures. v2: updated author list (added Yang Liu and Yuxin Li; marked core contributors and project lead) and added a release-date note on the first page
机构
*
University of Oklahoma(俄克拉荷马大学)
;
Imperial College London(伦敦帝国学院)
;
University of Michigan(密歇根大学)
;
Tencent(腾讯)
;
University of Edinburgh(爱丁堡大学)
;
University of Southern California(南加州大学)
;
Beijing Normal University(北京师范大学)
;
Beijing Normal-Hong Kong Baptist University(北京师范大学-香港浸会大学联合国际学院)
Cross-View Action Consistency for Camera-Robust Vision-Language-Action Policies
面向相机鲁棒性的视觉-语言-动作策略的跨视图动作一致性
Bingqi Huang, Bingchuan Wei, Xuan Wang, Yingkai Cai, Zhaokui Wang
机构
*
Tsinghua University(清华大学)
;
Informatics Institute, University of Amsterdam(阿姆斯特丹大学信息学研究所)
;
Faculty of Science, Vrije Universiteit Amsterdam(阿姆斯特丹自由大学理学院)
Grad-ECLIP: Gradient-based Visual and Textual Explanations for CLIP
Grad-ECLIP: 基于梯度的CLIP视觉与文本解释
Chenyang Zhao, Kun Wang, Janet H. Hsiao, Antoni B. Chan
机构
*
Department of Computer Science, City University of Hong Kong(香港城市大学计算机科学系)
;
Division of Social Science and Department of Computer Science & Engineering, Hong Kong University of Science & Technology(香港科学与技术大学社会科学学院及计算机科学与工程系)
;
SenseTime Group Ltd(时光集团有限公司)
Journal refZhao C, Wang K, Hsiao J H, et al. Grad-eclip: Gradient-based visual and textual explanations for clip[J]. IEEE Transactions on Pattern Analysis and Machine Intelligence, 2026