arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

Image Fusion

围绕 image fusion 的图像融合方法、数据集、评测与应用。

共收录 24
2608.13045 2026-08-14 cs.CV 新提交 78%

P2Fusion: Prompt-based Progressive Infrared-Visible Image Fusion via Dual-Prior Distillation

P2Fusion:基于提示的双先验蒸馏红外-可见光图像渐进融合

Yi Shi, Huichao Xie, Yuqing Wang, Mingyu Wang, Kaihui Yang, Yu Liu, Ruitao Lu, Lizhe Li, Junwei Han, Dingwen Zhang

机构 * Northwestern Polytechnical University(西北工业大学) Hefei University of Technology(合肥工业大学) Rocket Force University of Engineering(火箭军工程大学) Chongqing University of Posts and Telecommunications(重庆邮电大学)

专题命中 Image Fusion :image fusion(title,abstract)

AI总结 针对红外-可见光图像融合的信息差异挑战,提出基于双固有提示的P2Fusion框架,结合Teach-to-Fuse机制与GDER模块,在多数据集上实现SOTA性能并提升下游感知鲁棒性。

Comments Accepted by ECCV 2026. Website: https://p2fusion.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.03252 2026-08-05 cs.CV 新提交 78%

Clarity Contrast and Similarity Selection for Multi-Focus Image Fusion

用于多聚焦图像融合的清晰度对比与相似性选择

Yicheng Zhang, Haoyou Deng, Zhiqiang Li, Wenti Yin, Nong Sang, Changxin Gao

专题命中 Image Fusion :image fusion(title,abstract)

AI总结 该研究针对多聚焦图像融合中源图像交互不足的问题,提出CSNet模型,通过CCAM和相似性选择策略实现信息交互,在定量和定性评估中达到当前最优性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.00530 2026-08-04 cs.CV cs.AI 新提交 78%

Unleashing the Power of Text: Text-Guided Flow Matching for Image Fusion under Complex Degradations

释放文本的力量:面向复杂退化场景的文本引导流匹配图像融合

Axi Niu, Jieheng Li, Kang Zhang, Qingsen Yan, Jinqiu Sun, Yanning Zhang

专题命中 Image Fusion :image fusion(title,abstract)

AI总结 针对复杂退化下红外-可见光图像融合的挑战,提出文本引导的TGFusion框架,通过多流联合流Transformer实现动态文本引导,在多类退化场景下取得优异融合性能。

Comments 12 pages, 9 figures, including supplementary material

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.28565 2026-07-31 cs.CV 新提交 78%

MIND: Multimodal Intent-Driven Network via Diffusion Transformers for Medical Image Fusion

MIND:基于扩散Transformer的多模态意图驱动网络用于医学图像融合

Yunzhan Fu, Xiangyu Shen, Yifei Sun, Yuhan Chen, Jian Wu, Hongxia Xu

机构 * Transvascular Implantation Devices Research Institute, Zhejiang University(浙江大学血管内植入器械研究院) Zhejiang University(浙江大学) Hangzhou Institute of Technology, Xidian University(西安电子科技大学杭州研究院) Hangzhou Dianzi University(杭州电子科技大学)

专题命中 Image Fusion :image fusion(title,abstract)

AI总结 该研究针对现有医学图像融合方法缺乏对诊断意图深度理解的问题,提出基于DiTs的MIND网络,通过BioMedGPT、多尺度潜在适配器和医学语义一致性损失优化,在多数据集上取得优异效果,可提升脑肿瘤分割精度并支持交互式融合。

Comments 14pages, 14 figures, accepted by ACM MM2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.25338 2026-07-29 cs.AI 新提交 78%

Dual-Domain Manifold Modeling for Hyperspectral Image Fusion

用于高光谱图像融合的双域流形建模

Chengxin Xie, Qiya Song, Yangbangyan Jiang, Renwei Dian, Xudong Kang

机构 * College of Information Science and Engineering, Hunan Normal University(湖南师范大学信息科学与工程学院) Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所) School of Artificial Intelligence and Robotics, Hunan University(湖南大学人工智能与机器人学院)

专题命中 Image Fusion :image fusion(title,abstract)

AI总结 针对高光谱图像融合中几何约束建模难题,提出双域流形建模框架,通过拓扑感知Transformer和频率解耦的空间 - 光谱协同融合模块,有效整合光谱与空间信息,在多数据集实验中表现优于现有方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.23600 2026-07-28 cs.CV 新提交 78%

ConFusion: Continuous Fusion Space Learning for Fine-Grained Controllable Infrared and Visible Image Fusion

ConFusion:用于细粒度可控红外与可见光图像融合的连续融合空间学习

Guo Yurong, He Yufei, Li Yonghao, Chang Dongliang, Zhang Ke, Ma Zhanyu

机构 * North China Electric Power University(华北电力大学) Beijing University of Posts and Telecommunications(北京邮电大学)

专题命中 Image Fusion :image fusion(title,abstract)

AI总结 研究可控红外与可见光图像融合,提出ConFusion框架,通过高斯条件空间感知调制学习连续融合空间,采用双分支架构及相关模块实现实例级细粒度可控融合,实验证明其在融合质量和下游任务上达先进水平。

Comments 10 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.19879 2026-07-23 cs.CV 新提交 78%

Current Injection Spiking Neural Network for Infrared and Visible Image Fusion

用于红外与可见光图像融合的电流注入脉冲神经网络

Rui Zhao, Zhuoyuan Li, Wenrui Li, Yanchen Dong, Yajing Zheng, Giuseppe Valenzise, Weisi Lin

机构 * College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算与数据科学学院) The Hong Kong Polytechnic University(香港理工大学) Harbin Institute of Technology(哈尔滨工业大学) State Key Laboratory for Multimedia Information Processing, School of Computer Science, Peking University(北京大学计算机科学学院多媒体信息处理国家重点实验室)

专题命中 Image Fusion :image fusion(title,abstract)

AI总结 研究红外与可见光图像融合问题,提出CIS - Fuse脉冲网络,通过电流注入脉冲算子在膜电位水平实现跨模态融合,构建双向跨模态融合模块并部署在双分支架构上,实验表明其融合质量与基于ANN的方法相当且更节能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.03643 2026-07-07 eess.IV cs.CV 新提交 78%

Model Confidence-Guided Multi-Image Fusion of Fundus Images for Diabetic Retinopathy Diagnosis

模型置信度引导的多图像融合眼底图像糖尿病视网膜病变诊断方法

Ananya Raghu, Anisha Raghu, Alice S. Tang, Yannis M. Paulus, Tyson N. Kim, Tomiko T. Oskotsky

机构 * Massachusetts Institute of Technology(麻省理工学院) Wilmer Eye Institute, Department of Ophthalmology, Johns Hopkins University(约翰霍普金斯大学威尔默眼科研究所) Department of Biomedical Engineering, Johns Hopkins University(约翰霍普金斯大学生物医学工程系) Bakar Computational Health Sciences Institute, University of California San Francisco(加州大学旧金山分校巴卡计算健康科学研究所) Department of Ophthalmology, University of California San Francisco(加州大学旧金山分校眼科系) Division of Clinical Informatics and Digital Transformation, University of California San Francisco(加州大学旧金山分校临床信息学与数字化转型 division)

专题命中 Image Fusion :image fusion(title,abstract)

AI总结 针对资源有限地区糖网病早期筛查需求,提出置信度引导的多图像融合框架,整合多眼底视图提升诊断置信度与准确率,性能优于传统图像质量级联流水线,适配低延迟移动筛查场景。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.02572 2026-07-07 cs.CV cs.AI 新提交 78%

Additive Causal Construction for Transferable and Reconfigurable Cross-System Learning in Multi-Source Image Fusion

多源图像融合中用于可转移和可重构跨系统学习的加性因果构建

Zhizhong Fu, Wei Zhou, Zhaoyang Jiang, Yulong Lin, Yifu Hou, Xiaorong Ding, Qiang Yan, Yifan Chen

机构 * School of Life Science and Technology, University of Electronic Science and Technology of China(电子科技大学生命科学与技术学院) School of Health and Wellbeing, University of Glasgow(格拉斯哥大学健康与幸福学院) Department of Organ Transplantation, Sichuan Provincial People’s Hospital, University of Electronic Science and Technology of China(电子科技大学附属四川省人民医院器官移植科) School of Information and Software Engineering, University of Electronic Science and Technology of China(电子科技大学信息与软件工程学院) Department of Radiology, Huzhou Maternity & Child Health Care Hospital(湖州市妇幼保健院放射科) Hepatological Surgery Department, Huzhou Central Hospital, Fifth School of Clinical Medicine of Zhejiang Chinese Medical University(浙江中医药大学附属湖州中医院第五临床医学院湖州市中心医院肝外科)

专题命中 Image Fusion :image fusion(title,abstract)

AI总结 针对多源图像融合中跨系统差异和纠缠问题,提出加性因果构建框架,通过干预一致性建立共享因果‘锚’实现因果图可转移,将融合过程形式化为因果构建并量化不确定性确保可重构,改进因果表示学习。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.31242 2026-07-01 cs.CV 新提交 78%

UHD-MFF: Shattering Barriers in Multi-Focus Ultra-High-Definition Image Fusion via Learnable Lookup Tables

UHD-MFF:通过可学习查找表打破多焦点超高清图像融合的障碍

Yibing Zhang, Xunpeng Yi, Qinglong Yan, Yeda Wang, Han Xu, Jiayi Ma

机构 * Electronic Information School, Wuhan University(武汉大学电子信息学院) School of Robotics, Wuhan University(武汉大学机器人学院) School of Automation, Southeast University(东南大学自动化学院)

专题命中 Image Fusion :image fusion(title,abstract)

AI总结 针对超高清多焦点图像融合中的数据、模型和部署三大障碍,提出首个大规模超高清数据集UHD-MFF和基于可学习查找表的UMF-LUT框架,实现实时4K融合。

Comments Accepted by ECCV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.26812 2026-06-26 cs.CV 新提交 78%

Multi-modality Image Fusion under Adverse Weather: Mask-Guided Feature Restoration and Interaction

恶劣天气下的多模态图像融合:掩码引导的特征恢复与交互

Xilai Li, Xiaosong Li, Haishu Tan, Tao Ye, Huafeng Li, Hongbin Wang

机构 * Foshan University(佛山大学) China University of Mining and Technology, Beijing(中国矿业大学(北京)) Kunming University of Science and Technology(昆明理工大学)

专题命中 Image Fusion :image fusion(title,abstract)

AI总结 提出一种掩码引导的多模态图像融合方法,通过伪真实标签和掩码生成机制同时实现特征恢复与跨模态交互,在合成和真实数据集上超越现有方法。

Comments Accepted at ECCV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.12303 2026-06-11 cs.CV 新提交 78%

From 2D Grids to 1D Tokens: Reforming Shared Representations for Multimodal Image Fusion

从二维网格到一维标记:重塑多模态图像融合的共享表示

Yuchen Xian, Yunqiu Xu, Yang He, Yi Yang

机构 * University of Science and Technology of China(中国科学技术大学)

专题命中 Image Fusion :image fusion(title,abstract)

AI总结 提出基于冻结预训练图像标记器的紧凑一维标记接口,通过选择性标记编辑(STE)稀疏更新关键标记,在保持融合骨干网络不变的同时引导全局外观一致性,实现全局连贯与局部保真的最佳平衡。

Comments Accepted at the 43rd International Conference on Machine Learning (ICML 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.07985 2026-06-09 cs.CV cs.CL 新提交 78%

FMRFusion: Frequency-Aware Multi-View Representation Learning for Heterogeneous Image Fusion

FMRFusion: 面向异质图像融合的频率感知多视图表示学习

Tao Zhoua, Yunlong Liu, Qinghui Chen, Zekai Zhang, Minlong Sun, Changlin Biana, Dagang Li, Wenmin Wang, Jinglin Zhang

机构 * Shandong University(山东大学) Macau University of Science and Technology(澳门科技大学)

专题命中 Image Fusion :image fusion(title,abstract)

AI总结 提出FMRFusion网络,通过多尺度结构感知模块、双线性频率分解和跨视图互补交互,结合流匹配优化,实现红外与可见光图像融合,在夜间场景表现优异。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.03107 2026-08-05 cs.CV 新提交 71%

A Unified Resolution-Conditioned Framework for Orthogonal Line-Scanning Image Fusion

一种用于正交线扫描图像融合的统一分辨率条件框架

Yiming Gong, Kai Wang

专题命中 Image Fusion :image fusion(title)

AI总结 针对正交线扫描图像融合,提出基于RELA和FiLM的统一分辨率条件框架,可跨狭缝宽度适配,性能优于现有方法,能平滑泛化到未见过的中间配置。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.23136 2026-07-28 eess.IV cs.PF 新提交 71%

Optimized Embedded Implementation of Hyperspectral-Multispectral Image Fusion on Raspberry Pi

树莓派上高光谱-多光谱图像融合的优化嵌入式实现

Salah Eddine Brezini, Okba Bekhelifi, Oussama Mezouar, Chams Eddine Choucha, Sarra Boukhacheba, Fethi Abdelatif Dali

专题命中 Image Fusion :image fusion(title)

AI总结 针对高光谱图像数据量大难实时处理的问题,提出将密集运算迁移到特定框架结合XNNPACK后端进行优化的方法,在树莓派5平台部署,显著减少计算时间并保留融合质量,适用于相关遥感应用。

Comments paper is accepted for presentation at EDiS'2026: IEEE 5th International Conference on Embedded and Distributed Systems, Oran, Algeria, November 2-5, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.07848 2026-08-11 cs.CV 新提交 50%

IRPol-Fuse: Energy-structure coordination for infrared polarization fusion under low visibility

IRPol-Fuse:低能见度下红外偏振图像的能量-结构协同融合方法

Zhuangfan Huang, Chusheng Fang, Xiaosong Li, Yang Liua, Xiaoqi Cheng, Haishu Tan

机构 * Foshan University(佛山大学)

专题命中 Image Fusion :image fusion(abstract)

AI总结 该研究针对低能见度下红外偏振图像融合的缺陷,提出IRPol-Fuse框架,构建LI-PI数据集,实验验证其在目标与细节保留及下游任务中的优异性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.22773 2026-07-28 eess.IV cs.CV physics.med-ph 新提交 50%

Metric Surface Reconstruction of Neurosurgical Scenes from Monocular Operating Microscope Images and Microscope Pose

从单目手术显微镜图像和显微镜位姿重建神经外科场景的度量曲面

Thomas Bucher, Didier Neuenschwander, Thomas Petutschnigg, Michael Murek, David Bervini, Andreas Raabe, Manuela Eugster

专题命中 Image Fusion :image fusion(abstract)

AI总结 研究能否从单目手术显微镜图像和位姿数据重建神经外科场景的三维几何度量,利用预训练模型估计深度、泊松曲面重建点云成网格,结果显示该方法在模型设置中有技术可行性,支持相关手术技术进一步发展。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.09351 2026-07-13 cs.CV 新提交 50%

Simon-SR: Spatially Adaptive Modulation and Visual Prompt Adaptation for Text-Reinforced Super-Resolution

Simon-SR:用于文本增强超分辨率的空间自适应调制和视觉提示适配

Haotong Cheng, Yuxuan Li, Zijie Cui, Rongling Tan, Chenyuan Wang

机构 * College of Electronic Science and Engineering, Jilin University(吉林大学电子科学与工程学院)

专题命中 Image Fusion :image fusion(abstract)

AI总结 针对单图像超分辨率问题,提出Simon-SR框架,利用可学习提示进行语义挖掘和文本-图像融合,结合对比提示学习与空间自适应细化,实验证明该方法超越现有技术,在多项指标上有明显提升。

Comments Multi-modal Single Image Super-Resolution

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.02394 2026-07-03 stat.AP stat.ME 新提交 50%

Masked complex non-decimated wavelet features for patient-level classification of contrast-enhanced mammography

掩蔽复值非抽取小波特征用于对比增强乳腺X线摄影的患者级分类

Sara Antonijevic, Brani Vidakovic

专题命中 Image Fusion :image fusion(abstract)

AI总结 针对对比增强乳腺X线摄影(CESM)图像分类中图像类型信号可比性和患者多图像融合问题,提出掩蔽复值非抽取小波特征结合弹性网逻辑回归,在无泄漏评估下两种图像类型在患者级AUC上无统计差异,且特征可解释。

Comments 29 pages, 9 figures. Code available at https://github.com/saraantonijevic/Masked_Mammograms

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.00595 2026-07-03 cs.CV 新提交 50%

GADA: Geometry-Aware Deformable Aggregation for Image-Based Gaussian Splatting

GADA: 基于图像的高斯泼溅的几何感知可变形聚合

Siwoo Lim, Sunjae Yoon, Gwanhyeong Koo, Chang D. Yoo

机构 * Korea Advanced Institute of Science and Technology (KAIST)(韩国科学技术院) Chung-Ang University(中央大学)

专题命中 Image Fusion :image fusion(abstract)

AI总结 提出几何感知可变形聚合(GADA),通过可变形偏移迭代校正空间错位,并引入隐式置信加权机制抑制不可靠证据,在保持高频细节的同时实现2.13倍FPS提升。

Comments ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.20689 2026-06-23 cs.CV cs.LG 新提交 50%

NeoJaundice-AI: Smartphone-Based Neonatal Jaundice Detection Using Dual-Input Deep Learning and Synthetic Augmentation

NeoJaundice-AI: 基于智能手机的新生儿黄疸检测,采用双输入深度学习与合成增强

Rahul Patel, Nirjala Jarpula

机构 * Indian Institute of Information Technology Surat(印度信息技术学院苏拉特分校)

专题命中 Image Fusion :image fusion(abstract)

AI总结 提出NeoJaundice-AI系统,通过双分支EfficientNet-B0处理皮肤和巩膜图像,融合手工YCbCr颜色特征,实现四类严重度分类和胆红素回归,采用合成黄疸生成和肤色归一化,在印度新生儿数据集上达到91.8%准确率。

Comments 7 pages, 10 figures, 8 tables. IEEE conference format

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.15104 2026-06-16 cs.CV 新提交 50%

Text-Driven Fusion for Infrared and Visible Images: Achieving Image Scene Adaptation on Hyperbolic Space

红外与可见光图像的文本驱动融合:在双曲空间实现图像场景自适应

Huan Kang, Hui Li, Tianyang Xu, Tao Zhou, Xiao-Jun Wu, Josef Kittler

机构 * University of Science and Technology of China(中国科学技术大学)

专题命中 Image Fusion :image fusion(abstract)

AI总结 提出一种文本驱动的红外与可见光图像融合框架,利用双曲流形学习嵌入层次语义,通过BLIP文本提示引导视觉-属性对齐,实现无文本输入的自适应融合,性能优于现有方法。

Comments 14 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.08065 2026-06-09 physics.med-ph 新提交 50%

The Role of Free-breathing GRASP MRI in Accurate Phase Matching with 4D-CT for Motion Representation in Liver Cancer Radiotherapy

自由呼吸GRASP MRI在肝癌放疗中与4D-CT精确相位匹配以表征运动中的作用

Junchao Li, Shengqi Chen, Guohua Wu, Jianrong Dai, Jiayun Chen, Fei Liu

专题命中 Image Fusion :image fusion(abstract)

AI总结 研究自由呼吸GRASP MRI能否代表肝癌立体定向放疗中的呼吸运动,发现其仅在30%-60%呼吸相位(最佳50%)准确表征运动,需结合4D-CT或动态成像。

Comments 25 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.07651 2026-06-09 cs.LG cs.CV 新提交 50%

KITE: A Tri-Modal Transformer Integrating Text, Images, and Knowledge Graphs for Fake News Detection

KITE:一种融合文本、图像和知识图谱的三模态假新闻检测Transformer

Kevin Patel, Shashi Bhushan Jha

机构 * Department of Computer Science, University of West Florida(威斯福大学计算机科学系)

专题命中 Image Fusion :image fusion(abstract)

AI总结 提出三模态假新闻检测框架KITE,联合建模文本、视觉和知识表示,利用跨模态注意力整合特征,在基准数据集上显著优于单双模态基线。

详情

展开后加载摘要…

URL PDF HTML 收藏