VLC Fusion: Vision-Language Conditioned Sensor Fusion for Robust Object Detection
VLC Fusion:面向鲁棒目标检测的视觉-语言条件传感器融合
Aditya Taparia, Noel Ngu, Mario Leiva, Joshua Shay Kricheli, John Corcoran, Nathaniel D. Bastian, Gerardo Simari, Paulo Shakarian, Ransalu Senanayake
机构
*
Arizona State University(亚利桑那州立大学)
;
Department of Computer Science and Engineering, Universidad Nacional del Sur and Institute for Computer Science and Engineering(计算机科学与工程系,国家南方大学和计算机科学与工程研究所)
;
U.S. Department of Defense(美国国防部)
;
United States Military Academy(美国军事学院)
;
Syracuse University(雪城大学)
CommentsAccepted by the Proceedings of the 41st IEEE/ACM International Conference on Automated Software Engineering (ASE '26), 2026. This is the authors' version of the work; the definitive Version of Record is forthcoming
机构
*
School of Computer Science and Engineering(计算机科学与工程学院)
;
UNSW Sydney(新南威尔士大学悉尼分校)
;
School of Information and Communication Technology(信息与通信技术学院)
;
Griffith University(格里菲斯大学)
机构
*
Applied Artificial Intelligence and Intelligent Systems (AAIINS) Laboratory(应用人工智能与智能系统实验室)
;
Department of Computer Science and Engineering(计算机科学与工程系)
;
Department of Data Science and Artificial Intelligence(数据科学与人工智能系)
;
Department of Software Systems & Cybersecurity(软件系统与网络安全系)
;
Energy and Resources Institute, Faculty of Science and Technology(能源与资源研究所,科学与技术学院)
;
Faculty of Science and Technology(科学与技术学院)
;
School of Engineering and Energy(工程与能源学院)
机构
*
Jiangsu Key Laboratory of Networked Collective Intelligence, School of Mathematics, Southeast University(江苏网络集体智能重点实验室,数学学院,东南大学)
;
Jiangsu Key Laboratory of Networked Collective Intelligence, School of Cyber Science and Engineering, Southeast University(江苏网络集体智能重点实验室,网络科学与工程学院,东南大学)
;
State Key Laboratory of Mathematical Sciences, Academy of Mathematics and Systems Science, University of Chinese Academy of Sciences(数学科学国家重点实验室,数学与系统科学研究院,中国科学院大学)
CommentsThis paper has been accepted for presentation at INTERSPEECH 2026 and as non-archival paper at ICML 2026 Workshop on Machine Learning for Audio