Jiageng Wen, Shengjie Zhao, Bing Li, Jiafeng Huang, Kenan Ye, Hao Deng
机构
*
Shanghai Research Institute for Intelligent Autonomous Systems, Tongji University(同济大学智能自主系统上海研究院)
;
School of Computer Science and Technology, Tongji University(同济大学计算机科学与技术学院)
;
School of Mechatronic Engineering and Automation, Shanghai University(上海大学机械电子工程与自动化学院)
CommentsAccepted to ICLR 2026. This arXiv version includes an additional appendix (Appendix 15) containing further philosophical discussion not included in the official ICLR peer-reviewed version
机构
*
School of Artificial Intelligence and Computer Science, Jiangnan University, Wuxi, China(江南大学人工智能与计算机科学学院)
;
School of Electronic and Information Engineering, Suzhou University of Science and Technology, Suzhou, China(苏州科技大学电子与信息工程学院)
专题命中
红外-可见光融合
:image fusion(title,abstract);infrared and visible(title,abstract);分类 cs.CV
ITO: Images and Texts as One via Synergizing Multiple Alignment and Training-Time Fusion
通过协同多模态对齐和训练时融合实现图像与文本一体化:ITO
Hanpeng Liu, Yaqian Li, Zidan Wang, Shuoxi Zhang, Zonglin Zhao, Zihao Bo, Rinyoichi Takezoe, Kaiwen Long, Kun He
机构
*
School of Computer Science(计算机科学学院)
;
Huazhong University of Science and Technology(华中科技大学)
;
Li Auto Inc.(力汽车公司)
;
Institute of AI for Industries, Chinese Academy of Sciences(产业人工智能研究院,中国科学院)
Toward Multimodal Industrial Fault Analysis: A Single-Speed Chain Conveyor Dataset with Audio and Vibration Signals
迈向多模态工业故障分析:一个包含音频和振动信号的单速链式输送机数据集
Zhang Chen, Yucong Zhang, Xiaoxiao Miao, Ming Li
机构
*
Digital Innovation Research Center, Duke Kunshan University(杜克昆山大学数字创新研究中心)
;
School of Artificial Intelligence, The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)人工智能学院)
;
School of Artificial Intelligence, Wuhan University(武汉大学人工智能学院)
;
School of Computer Science, Wuhan University(武汉大学计算机学院)