From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving
从表征互补性到双系统:协同VLM和纯视觉骨干网络用于端到端驾驶
Sining Ang, Yuguang Yang, Chenxu Dang, Canyu Chen, Cheng Chi, Haiyan Liu, Xuanyao Mao, Jason Bao, Xuliang, Bingchuan Sun, Yan Wang
机构
*
Department of Automation, University of Science and Technology of China(中国科学技术大学自动化系)
;
School of Electronic Information Engineering, Beihang University(北京航空航天大学电子信息工程学院)
;
School of Artificial Intelligence and Automation, Huazhong University of Science and Technology(华中科技大学人工智能与自动化学院)
;
National Superior College for Engineers, Beihang University(北京航空航天大学国家级工程师学院)
;
Beijing Academy of Artificial Intelligence(北京人工智能研究院)
;
Lenovo Group Limited(联想集团有限公司)
;
Institute for AI Industry Research, Tsinghua University(清华大学人工智能产业研究院)
Comments14 pages, Due to arXiv's abstract length limit, the web version of the abstract has been shortened while preserving the paper's original scope, claims, and results. The implementation and experimental artifacts are publicly available at https://irish-kw.github.io/TGCM_Website/
Semantic Evidence Regulation via Relational Bias for Zero-Shot Object Navigation
通过关系归纳偏差重新思考具身导航
Weitao An, Chenghao Xu, Xu Yang, Cheng Deng
机构
*
School of Electronic Engineering, Xidian University(西安电子科技大学电子工程学院)
;
School of Information Science and Engineering, Hohai University(河海大学信息科学与工程学院)
Breaking Déjà Vu: Independent Auditing of Visual Place Recognition through Vision-Language Reasoning
打破似曾相识:通过视觉语言推理对视觉场所识别进行独立审计
Sania Waheed, Michael Milford, Sarvapali D. Ramchurn, Shoaib Ehsan
机构
*
School of Electronics and Computer Science, University of Southampton(南安普顿大学电子与计算机科学学院)
;
School of Electrical Engineering and Computer Science, Queensland University of Technology(昆士兰科技大学电气工程与计算机科学学院)
;
School of Computer Science and Electronic Engineering, University of Essex(埃塞克斯大学计算机科学与电子工程学院)
机构
*
Nanyang Technological University(南洋理工大学)
;
Centre for Frontier AI Research, A*STAR(A*STAR前沿人工智能研究中心)
;
Allen Institute for AI(艾伦人工智能研究所)
;
University of Washington(华盛顿大学)
机构
*
Graduate School of Information Science and Technology, Hokkaido University(北海道大学信息科学研究生院)
;
Graduate School of Computer Science, George Mason University(乔治·马歇尔大学计算机科学研究生院)
机构
*
Seoul National University(首尔大学)
;
Robotics Lab, Hyundai Motor Company(现代汽车公司机器人实验室)
;
Pohang University of Science and Technology (POSTECH)(浦项科技大学)
Reasoning in machine vision by learning fast and slow thinking
通过快速与慢速思考学习进行机器视觉推理
Shaheer U. Saeed, Yipei Wang, Veeru Kasivisvanathan, Brian R. Davidson, Matthew J. Clarkson, Yipeng Hu, Daniel C. Alexander
机构
*
Centre for Bioengineering, Queen Mary University of London(伦敦玛丽女王大学生物工程中心)
;
Digital Environment Research Institute, Queen Mary University of London(伦敦玛丽女王大学数字环境研究所)
;
School of Engineering and Materials Science, Queen Mary University of London(伦敦玛丽女王大学工程与材料科学学院)
;
UCL Hawkes Institute, University College London(伦敦大学学院UCL霍克斯研究所)
;
Department of Medical Physics and Biomedical Engineering, University College London(伦敦大学学院医学物理与生物医学工程系)
;
Centre for Urology Imaging, Prostate, AI and Surgical Studies (COMPASS) Research Group, Division of Surgery and Interventional Science, University College London(伦敦大学学院外科与介入科学部泌尿影像、前列腺、人工智能与手术研究(COMPASS)组中心)
;
Department of Urology, Comprehensive Cancer Center, Medical University of Vienna(维也纳医科大学综合癌症中心泌尿科)
;
Division of Surgery and Interventional Science, University College London(伦敦大学学院外科与介入科学部)
;
Department of Computer Science, University College London(伦敦大学学院计算机科学系)