arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

多模态信息融合

面向图像、视频、多传感器和跨模态感知的信息融合,包括 Image Fusion、红外可见光、遥感、医学影像、LiDAR/雷达/相机和音视频融合。

共收录 4740 信号源:cs.CV, eess.IV, eess.SP, cs.RO, cs.MM

1. 红外-可见光融合 533 篇

2603.04745 2026-03-06 cs.CV 57%

Toward Real-world Infrared Image Super-Resolution: A Unified Autoregressive Framework and Benchmark Dataset

迈向真实世界红外图像超分辨率:一个统一的自回归框架和基准数据集

Yang Zou, Jun Ma, Zhidong Jiao, Xingyuan Li, Zhiying Jiang, Jinyuan Liu

机构 * School of Computer Science, Northwestern Polytechnical University(西北工业大学计算机学院) School of Software Technology & DUT-RU International School of ISE, Dalian University of Technology(大连理工大学软件学院及DUT-RU国际信息工程学院) School of Computer Science, Zhejiang University(浙江大学计算机学院) College of Information Science and Technology, Dalian Maritime University(大连海事大学信息科学与技术学院)

专题命中 红外-可见光融合 :infrared and visible(abstract);分类 cs.CV

AI总结 本文提出Real-IISR框架和FLIR-IISR数据集,通过热-结构引导的自回归方法提升真实世界红外图像超分辨率性能。

Comments This paper was accepted by CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.16654 2026-02-27 cs.CV 57%

IV-tuning: Parameter-Efficient Transfer Learning for Infrared-Visible Tasks

IV-tuning: 用于红外-可见任务的参数高效迁移学习

Yaming Zhang, Chenqiang Gao, Fangcen Liu, Junjie Guo, Lan Wang, Xinggan Peng, Deyu Meng

专题命中 红外-可见光融合 :infrared and visible(abstract);分类 cs.CV

AI总结 IV-tuning通过冻结参数和高效微调,提升红外-可见任务的泛化能力和计算效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.06363 2026-02-09 cs.CV 57%

Robust Pedestrian Detection with Uncertain Modality

具有不确定模态的鲁棒行人检测

Qian Bie, Xiao Wang, Bin Yang, Zhixi Yu, Jun Chen, Xin Xu

专题命中 红外-可见光融合 :information fusion(abstract);分类 cs.CV

AI总结 本文提出AUNet,通过自适应不确定性感知网络在不确定输入下实现鲁棒行人检测,结合UMVR和MAI模块提升跨模态信息融合效果。

Comments Due to the limitation "The abstract field cannot be longer than 1,920 characters", the abstract here is shorter than that in the PDF file

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08977 2026-01-15 cs.CV 57%

Thermo-LIO: A Novel Multi-Sensor Integrated System for Structural Health Monitoring

Thermo-LIO:一种用于结构健康监测的新型多传感器集成系统

Chao Yang, Haoyuan Zheng, Yue Ma

专题命中 红外-可见光融合 :multimodal fusion(abstract);分类 cs.CV

AI总结 Thermo-LIO 通过融合热成像与高分辨率 LiDAR,提升大规模结构健康监测的精度和覆盖范围。

Comments 27pages,12figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06400 2026-01-14 cs.CV 57%

Perceptual Region-Driven Infrared-Visible Co-Fusion for Extreme Scene Enhancement

感知驱动的红外可见联合融合用于极端场景增强

Jing Tao, Yonghong Zong, Banglei Guan, Pengju Sun, Taihang Lei, Yang Shanga, Qifeng Yu

机构 * College of Aerospace Science and Engineering, National University of Defense Technology(航天科学与工程学院,国防科技大学) Hunan Provincial Key Laboratory of Image Measurement and Vision Navigation(湖南省图像测量与视觉导航重点实验室) Beijing Institute of Tracking and Telecommunication Technology(北京跟踪与电信技术研究所) National Key Laboratory of Space Integrated Information System(空间一体化信息系统国家重点实验室)

专题命中 红外-可见光融合 :multi-exposure(abstract);分类 cs.CV

AI总结 本文提出一种基于区域感知的红外可见联合融合框架,通过多曝光和多模态成像技术,在极端条件下提升图像清晰度和融合性能。

Comments The paper has been accepted and officially published by OPTICS AND LASER TECHNOLOGY

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.02249 2026-01-06 cs.CV 57%

SLGNet: Synergizing Structural Priors and Language-Guided Modulation for Multimodal Object Detection

SLGNet: 结合结构先验与语言引导调制的多模态目标检测

Xiantai Xiang, Guangyao Zhou, Zixiao Wen, Wenshuai Li, Ben Niu, Feng Wang, Lijia Huang, Qiantong Wang, Yuhan Liu, Zongxu Pan, Yuxin Hu

机构 * Aerospace Information Research Institute, Chinese Academy of Sciences(中国科学院航天信息研究所) Key Laboratory of Target Cognition and Application Technology, Chinese Academy of Sciences(中国科学院目标认知与应用技术重点实验室) School of Electronic, Electrical and Communication Engineering, University of Chinese Academy of Sciences(中国科学院大学电子电气与通信工程学院) School of Software Engineering, Xi’an Jiaotong University(西安交通大学软件学院)

专题命中 红外-可见光融合 :multimodal fusion(abstract);分类 cs.CV

AI总结 SLGNet通过结合结构先验与语言引导调制,在冻结的ViT基础上实现高效多模态目标检测,提升环境适应性和检测性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.10294 2026-01-06 cs.CV 57%

A Mutual-Structure Weighted Sub-Pixel Multimodal Optical Remote Sensing Image Matching Method

一种互结构加权的子像素多模光学遥感图像匹配方法

Tao Huang, Hongbo Pan, Nanxi Zhou, Siyuan Zou, Shun Zhou

专题命中 红外-可见光融合 :visible-infrared(abstract);分类 cs.CV

AI总结 本文提出了一种基于相位一致性互结构加权的子像素多模光学遥感图像匹配方法,通过粗到细的框架提升匹配精度,实测平均精度达0.4像素。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.23252 2026-01-06 cs.CV 57%

DGE-YOLO: Dual-Branch Gathering and Attention for Accurate UAV Object Detection

DGE-YOLO:双分支聚集与注意力用于准确的无人机目标检测

Kunwei Lv, Zhiren Xiao, Hang Ren, Ping Lan

专题命中 红外-可见光融合 :infrared and visible(abstract);分类 cs.CV

AI总结 DGE-YOLO 通过双分支架构和高效多尺度注意力机制,提升多模态无人机目标检测的准确性和鲁棒性。

Comments 5 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.24243 2026-01-01 cs.CV 57%

MambaSeg: Harnessing Mamba for Accurate and Efficient Image-Event Semantic Segmentation

MambaSeg: 利用Mamba实现准确且高效的图像事件语义分割

Fuqiang Gu, Yuanke Li, Xianlei Long, Kangping Ji, Chao Chen, Qingyi Gu, Zhenliang Ni

专题命中 红外-可见光融合 :multimodal fusion(abstract);分类 cs.CV

AI总结 MambaSeg通过双分支框架和双维交互模块,实现高效准确的多模态语义分割,适用于快速运动和低光条件下的图像事件分割任务。

Comments Accepted by AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.22896 2025-12-01 cs.CV 57%

DM$^3$T: Harmonizing Modalities via Diffusion for Multi-Object Tracking

DM$^3$T: 通过扩散和谐多模态以实现多目标跟踪

Weiran Li, Yeqiang Liu, Yijie Wei, Mina Han, Qiannan Guo, Zhenbo Li

机构 * China Agricultural University(中国农业大学) Beijing Normal University(北京师范大学)

专题命中 红外-可见光融合 :multimodal fusion(abstract);分类 cs.CV

AI总结 DM$^3$T通过扩散模型实现多模态特征对齐,提升多目标跟踪的准确性和鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06406 2025-11-11 cs.CV cs.AI 57%

On Modality Incomplete Infrared-Visible Object Detection: An Architecture Compatibility Perspective

Shuo Yang, Yinghui Xing, Shizhou Zhang, Zhilong Niu

机构 * Shuo Yang Yinghui Xing Shizhou Zhang Zhilong Niu

专题命中 红外-可见光融合 :infrared and visible(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03044 2025-11-11 cs.CV 57%

DCDB: Dynamic Conditional Dual Diffusion Bridge for Ill-posed Multi-Tasks

Chengjie Huang, Jiafeng Yan, Jing Li, Lu Bai

专题命中 红外-可见光融合 :visible-infrared(abstract);分类 cs.CV

Comments The article contains factual errors

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17078 2025-10-21 cs.CV 57%

Towards a Generalizable Fusion Architecture for Multimodal Object Detection

Jad Berjawi, Yoann Dupas, Christophe C'erin

机构 * Université Grenoble Alpes(格勒诺布尔大学) Université Sorbonne Paris Nord(巴黎-萨克勒大学) INRIA(法国国家信息与自动化研究所)

专题命中 红外-可见光融合 :multimodal fusion(abstract);分类 cs.CV

Comments 8 pages, 8 figures, accepted at ICCV 2025 MIRA Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.10445 2025-08-15 cs.CV 57%

DOD-SA: Infrared-Visible Decoupled Object Detection with Single-Modality Annotations

Hang Jin, Chenqiang Gao, Junjie Guo, Fangcen Liu, Kanghui Tian, Qinyao Chang

专题命中 红外-可见光融合 :infrared and visible(abstract);分类 cs.CV

Comments 9 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20574 2025-07-29 cs.CV 57%

LSFDNet: A Single-Stage Fusion and Detection Network for Ships Using SWIR and LWIR

Yanyin Guo, Runxuan An, Junwei Li, Zhiyuan Zhang

机构 * Zhejiang University(浙江大学) Singapore Management University(新加坡管理学院)

专题命中 红外-可见光融合 :image fusion(abstract);分类 cs.CV

Comments ACMMM2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17237 2025-07-17 cs.CV cs.AI 57%

Strong Baseline: Multi-UAV Tracking via YOLOv12 with BoT-SORT-ReID

Yu-Hsi Chen

机构 * The University of Melbourne(墨尔本大学)

专题命中 红外-可见光融合 :information fusion(abstract);分类 cs.CV

Comments 10 pages, 5 figures, 5 tables

Journal ref Proceedings of the Computer Vision and Pattern Recognition Conference (CVPR) Workshops, 2025, pp. 6573-6582

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.15137 2025-07-15 cs.CV 57%

Multispectral Detection Transformer with Infrared-Centric Feature Fusion

Seongmin Hwang, Daeyoung Han, Moongu Jeon

机构 * Artificial Intelligence Graduate School, Gwangju Institute of Science and Technology (GIST)(人工智能研究生院,全州科学技术院(GIST)) School of Electrical Engineering and Computer Science, Gwangju Institute of Science and Technology (GIST)(电气工程与计算机科学学院,全州科学技术院(GIST))

专题命中 红外-可见光融合 :sensor fusion(abstract);分类 cs.CV

Comments Under Review

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.20803 2025-06-25 cs.CV 57%

Two-Stream Spatial-Temporal Transformer Framework for Person Identification via Natural Conversational Keypoints

Masoumeh Chapariniya, Hossein Ranjbar, Teodora Vukovic, Sarah Ebling, Volker Dellwo

专题命中 红外-可见光融合 :feature-level fusion(abstract);分类 cs.CV

Comments I would like to withdraw this submission due to the need for substantial revisions in the results and analysis. I plan to correct and improve the study and submit a more complete version in the near future

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.12536 2025-06-17 cs.RO cs.AI 57%

Deep Fusion of Ultra-Low-Resolution Thermal Camera and Gyroscope Data for Lighting-Robust and Compute-Efficient Rotational Odometry

Farida Mohsen, Ali Safa

机构 * College of Science and Engineering, Hamad Bin Khalifa University(哈马德·本·哈利法大学科学与工程学院)

专题命中 红外-可见光融合 :sensor fusion(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.08953 2025-06-11 cs.CV 57%

Cross-Spectral Body Recognition with Side Information Embedding: Benchmarks on LLCM and Analyzing Range-Induced Occlusions on IJB-MDF

Anirudh Nanduri, Siyuan Huang, Rama Chellappa

机构 * University of Maryland College Park(马里兰大学 College Park 分校) Johns Hopkins University(约翰霍普金斯大学)

专题命中 红外-可见光融合 :visible-infrared(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.00262 2025-05-16 cs.RO 57%

CRADMap: Applied Distributed Volumetric Mapping with 5G-Connected Multi-Robots and 4D Radar Perception

Maaz Qureshi, Alexander Werner, Zhenan Liu, Amir Khajepour, George Shaker, William Melek

机构 * Faculty of Mechanical and Mechatronics Engineering, University of Waterloo(机械与机电工程学院,滑铁卢大学)

专题命中 红外-可见光融合 :sensor fusion(abstract);分类 cs.RO

Comments 7 pages, 5 figures, IEEE, ICARM

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.04526 2025-05-08 cs.CV cs.AI 57%

DFVO: Learning Darkness-free Visible and Infrared Image Disentanglement and Fusion All at Once

Qi Zhou, Yukai Shi, Xiaojun Yang, Xiaoyu Xian, Lunjia Liao, Ruimao Zhang, Liang Lin

机构 * School of Information Engineering, Guangdong University of Technology(广东技术大学信息工程学院) Key Laboratory of Photonic Technology for Integrated Sensing and Communication, Ministry of Education of China(中国教育部长 Photonic Technology for Integrated Sensing and Communication 重点实验室) CRRC Institute Co., Ltd.(CRRC研究院) UBTECH Robotics Co., Ltd(UBTECH机器人有限公司) School of Data Science, The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)数据科学学院) School of Data and Computer Science, Sun Yat-sen University(中山大学数据与计算机科学学院)

专题命中 红外-可见光融合 :image fusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.03565 2025-05-07 cs.RO cs.SY eess.SY 57%

Thermal-LiDAR Fusion for Robust Tunnel Localization in GNSS-Denied and Low-Visibility Conditions

Lukas Schichler, Karin Festl, Selim Solmaz, Daniel Watzenig

机构 * Virtual Vehicle Research GmbH(虚拟车辆研究有限公司)

专题命中 红外-可见光融合 :sensor fusion(abstract);分类 cs.RO

Comments Submitted to IAVVC 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.19244 2025-05-07 cs.CV 57%

Semantic-Aligned Learning with Collaborative Refinement for Unsupervised VI-ReID

De Cheng, Lingfeng He, Nannan Wang, Dingwen Zhang, Xinbo Gao

机构 * Xidian University(西电大学) Northwestern Polytechnical University(西北工业大学) Chongqing University of Posts and Telecommunications(重庆邮电大学)

专题命中 红外-可见光融合 :visible-infrared(abstract);分类 cs.CV

Comments Accepted by IJCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.07520 2025-04-30 cs.CV 57%

Instruct-ReID: A Multi-purpose Person Re-identification Task with Instructions

Weizhen He, Yiheng Deng, Shixiang Tang, Qihao Chen, Qingsong Xie, Yizhou Wang, Lei Bai, Feng Zhu, Rui Zhao, Wanli Ouyang, Donglian Qi, Yunfeng Yan

机构 * Zhejiang University(浙江大学) Shanghai AI Laboratory(上海人工智能实验室) SenseTime Research(商汤科技研究院) Liaoning Technical University(辽宁工程技术大学) Shanghai Jiao Tong University(上海交通大学) Qing Yuan Research Institute, Shanghai Jiao Tong University(上海交通大学清元研究院)

专题命中 红外-可见光融合 :visible-infrared(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.10331 2025-04-22 cs.CV 57%

LL-Gaussian: Low-Light Scene Reconstruction and Enhancement via Gaussian Splatting for Novel View Synthesis

Hao Sun, Fenggen Yu, Huiyao Xu, Tao Zhang, Changqing Zou

机构 * Zhejiang Lab(浙江实验室) University of Chinese Academy of Sciences(中国科学院大学) State Key Lab of CAD&CG, Zhejiang University(浙江大学计算机辅助设计与图形学国家重点实验室) Simon Fraser University(Simon Fraser大学) Hangzhou Dianzi University(杭州电子科技大学)

专题命中 红外-可见光融合 :multi-exposure(abstract);分类 cs.CV

Comments Project page: https://sunhao242.github.io/LL-Gaussian_web.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19654 2025-04-01 cs.CV cs.AI cs.LG 57%

RGB-Th-Bench: A Dense benchmark for Visual-Thermal Understanding of Vision Language Models

Mehdi Moshtaghi, Siavash H. Khajavi, Joni Pajarinen

专题命中 红外-可见光融合 :visible-infrared(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.05889 2025-03-21 cs.CV cs.AI cs.CL 57%

CREMA: Generalizable and Efficient Video-Language Reasoning via Multimodal Modular Fusion

Shoubin Yu, Jaehong Yoon, Mohit Bansal

专题命中 红外-可见光融合 :multimodal fusion(abstract);分类 cs.CV

Comments ICLR 2025; first two authors contributed equally. Project page: https://CREMA-VideoLLM.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.10931 2025-03-17 cs.CV 57%

Multi-Domain Biometric Recognition using Body Embeddings

Anirudh Nanduri, Siyuan Huang, Rama Chellappa

专题命中 红外-可见光融合 :visible-infrared(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.07249 2025-03-11 cs.CV 57%

Text-IRSTD: Leveraging Semantic Text to Promote Infrared Small Target Detection in Complex Scenes

Feng Huang, Shuyuan Zheng, Zhaobing Qiu, Huanxian Liu, Huanxin Bai, Liqiong Chen

专题命中 红外-可见光融合 :information fusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏