arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2026-03-11 至 2026-03-11 共收录 49 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 49 篇

2603.09931 2026-03-11 cs.CV cs.AI 86%

Adaptive Clinical-Aware Latent Diffusion for Multimodal Brain Image Generation and Missing Modality Imputation

自适应临床感知潜在扩散用于多模态脑图像生成与缺失模态插值

Rong Zhou, Houliang Zhou, Yao Su, Brian Y. Chen, Yu Zhang, Lifang He, Alzheimer's Disease Neuroimaging Initiative

专题命中 扩散模型 :diffusion(title,abstract);image generation(title);分类 cs.CV

AI总结 ACADiff通过自适应临床感知扩散模型,实现多模态脑图像生成与缺失模态插值,提升阿尔茨海默病诊断的准确性和鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.15328 2026-03-11 cs.LG cs.CV q-bio.NC 83%

Kuramoto Orientation Diffusion Models

Kuramoto方向扩散模型

Yue Song, T. Anderson Keller, Sevan Brodjian, Takeru Miyato, Yisong Yue, Pietro Perona, Max Welling

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

AI总结 Kuramoto方向扩散模型利用生物系统中的相位同步机制,通过周期域构建生成模型,实现对方向性图像的结构化生成。

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.19996 2026-03-11 cs.CV cs.AI 83%

DP-IQA: Utilizing Diffusion Prior for Blind Image Quality Assessment in the Wild

DP-IQA: 利用扩散先验进行野外盲图像质量评估

Honghao Fu, Yufei Wang, Wenhan Yang, Alex C. Kot, Bihan Wen

机构 * School of Electrical and Electronic Engineering, Nanyang Technological University(南洋理工大学电子与电气工程学院) PengCheng Laboratory(鹏城实验室)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

AI总结 DP-IQA利用预训练扩散模型的先验知识,通过轻量级CNN模型提升盲图像质量评估的性能和泛化能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.08928 2026-03-11 cs.CV 83%

TIDE: Text-Informed Dynamic Extrapolation with Step-Aware Temperature Control for Diffusion Transformers

TIDE: 文本引导的动态外推与步感知温度控制用于扩散变换器

Yihua Liu, Fanjiang Ye, Bowen Lin, Rongyu Fang, Chengming Zhang

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

AI总结 TIDE通过文本锚定机制和动态温度控制,实现无需训练的高分辨率文本到图像生成,有效解决注意力稀释导致的结构退化问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09819 2026-03-11 cs.CV 79%

ConfCtrl: Enabling Precise Camera Control in Video Diffusion via Confidence-Aware Interpolation

ConfCtrl: 通过置信度感知插值实现视频扩散中的精确相机控制

Liudi Yang, George Eskandar, Fengyi Shen, Mohammad Altillawi, Yang Bai, Chi Zhang, Ziyuan Liu, Abhinav Valada

机构 * University of Freiburg(弗赖堡大学) Ludwig Maximilian University of Munich(慕尼黑路德维希-马克西米利安大学) Technical University of Munich(慕尼黑技术大学) Huawei Heisenberg Research Center (Munich)(华为海森堡研究中心(慕尼黑))

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 ConfCtrl通过置信度感知插值实现视频扩散中的精确相机控制,解决大视角变化下新视角生成问题,提升几何一致性与视觉合理性。

Comments 13 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09657 2026-03-11 cs.CV cs.AI cs.ET eess.IV 79%

When to Lock Attention: Training-Free KV Control in Video Diffusion

何时锁定注意力:视频扩散中的无训练KV控制

Tianyi Zeng, Jincheng Gao, Tianyi Wang, Zijie Meng, Miao Zhang, Jun Yin, Haoyuan Sun, Junfeng Jiao, Christian Claudel, Junbo Tan, Xueqian Wang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 KV-Lock通过动态调度KV融合比和CFG尺度,提升视频扩散模型中前景质量与背景一致性的平衡,实现无训练的高效视频编辑。

Comments 18 pages, 9 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09408 2026-03-11 cs.CV cs.AI cs.LG 79%

Reviving ConvNeXt for Efficient Convolutional Diffusion Models

复兴 ConvNeXt 以实现高效的卷积扩散模型

Taesung Kwon, Lorenzo Bianchi, Lennart Wittke, Felix Watine, Fabio Carrara, Jong Chul Ye, Romann Weber, Vinicius Azevedo

机构 * KAIST(韩国科学技术院) ETH Zürich(苏黎世联邦理工学院) ISTI-CNR(意大利国家研究理事会信息与自动化研究所) University of Pisa(比萨大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 本文提出全卷积扩散模型 FCDM-XL,通过使用 ConvNeXt 结构实现高效扩散模型,以更少的 FLOPs 和训练步骤达到与 DiT-XL 相当的性能。

Comments CVPR 2026. Official implementation: https://github.com/star-kwon/FCDM

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27602 2026-03-11 cs.CV 79%

Who Made This? Fake Detection and Source Attribution with Diffusion Features

谁制造了这个?通过扩散特征进行伪造检测和来源归因

Simone Bonechi, Paolo Andreini, Barbara Toniella Corradini

机构 * Department of Information Engineering and Mathematics, University of Siena(信息工程与数学系,锡耶纳大学) AIGO, Italian Institute of Technology(AIGO,意大利理工学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 FRIDA通过分析扩散特征,利用k-最近邻方法和紧凑神经分类器实现伪造图像检测与来源归因,达到跨生成器检测的最优性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.24417 2026-03-11 cs.CV 79%

EasyText: Controllable Diffusion Transformer for Multilingual Text Rendering

EasyText: 用于多语言文本渲染的可控扩散变换器

Runnan Lu, Yuxuan Zhang, Jiaming Liu, Haofan Wang, Yiren Song

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 EasyText提出了一种基于DiT的可控扩散变换器,用于多语言文本渲染,通过字符位置编码和插值技术实现精确生成,并利用大规模数据集提升性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.01547 2026-03-11 cs.CV 79%

Semi-Supervised Biomedical Image Segmentation via Diffusion Models and Teacher-Student Co-Training

半监督生物医学图像分割:通过扩散模型和教师-学生协同训练

Luca Ciampi, Gabriele Lagani, Giuseppe Amato, Fabrizio Falchi

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 本文提出了一种基于扩散模型的半监督生物医学图像分割框架,通过教师-学生协同训练和多轮伪标签生成策略,在有限标注数据下提升分割性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09054 2026-03-11 cs.CV 79%

Spectral-Structured Diffusion for Single-Image Rain Removal

谱结构扩散用于单图像雨去除

Yucheng Xing, Xin Wang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 SpectralDiff通过引入结构化的频谱扰动,提升单图像雨去除的效率和性能。

Comments 15 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.08998 2026-03-11 cs.CV 79%

Diffusion-Based Authentication of Copy Detection Patterns: A Multimodal Framework with Printer Signature Conditioning

基于扩散的复制检测模式认证:一种多模态框架与打印机签名条件化

Bolutife Atoki, Iuliia Tkachenko, Bertrand Kerautret, Carlos Crispim-Junior

机构 * Université Lumière Lyon 2, CNRS, INSA Lyon, Universite Claude Bernard Lyon 1, LIRIS UMR5205(里摩日大学里昂2分校、法国国家科学研究中心、里昂国立应用科学学院、里昂大学克莱尔-贝尔纳分校、LIRIS UMR5205)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 本文提出基于扩散的多模态认证框架,利用打印机签名条件化技术,实现对复制检测模式的高效认证,优于传统方法和深度学习方法。

Comments Accepted at WACV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09184 2026-03-11 cs.LG cs.AI 78%

Latent-DARM: Bridging Discrete Diffusion And Autoregressive Models For Reasoning

Latent-DARM: 联结离散扩散与自回归模型以实现推理

Lina Berrayana, Ahmed Heakl, Abdullah Sohail, Thomas Hofmann, Salman Khan, Wei Chen

机构 * EPFL(苏黎世联邦理工学院) MBZUAI(马克斯·普朗克人工智能研究所) ETH Zürich(苏黎世联邦理工学院) Microsoft Research Asia(微软亚洲研究院)

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 Latent-DARM通过结合离散扩散模型与自回归模型,提升多智能体系统在推理任务中的表现和协作效率。

Comments Published at LIT Workshop at ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01367 2026-03-11 cs.LG 78%

DUEL: Exact Likelihood for Masked Diffusion via Deterministic Unmasking

DUEL:通过确定性解屏蔽实现掩码扩散的精确似然

Gilad Turok, Chris De Sa, Volodymyr Kuleshov

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 DUEL通过确定性解屏蔽实现MDMs的精确似然计算,首次提供正确困惑度,显著提升MDMs性能,揭示其在领域内和零样本基准上的优势。

Comments 22 pages, 5 figures 8 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.02552 2026-03-11 cs.CL 78%

From Veracity to Diffusion: Adressing Operational Challenges in Moving From Fake-News Detection to Information Disorders

从真实性到扩散:从假新闻检测到信息紊乱的运营挑战

Francesco Paolo Savatteri, Chahan Vidal-Gorène, Florian Cafiero

机构 * Ecole nationale des chartes - PSL(国家档案学院-PSL)

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本文探讨了从假新闻检测到信息紊乱的转变,通过比较两个数据集,分析预测目标变化对性能的影响,并提出轻量透明的管道方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06385 2026-03-11 cs.RO cs.SY eess.SY 78%

From Demonstrations to Safe Deployment: Path-Consistent Safety Filtering for Diffusion Policies

从演示到安全部署:扩散策略的路径一致安全过滤

Ralf Römer, Julian Balletshofer, Jakob Thumm, Marco Pavone, Angela P. Schoellig, Matthias Althoff

机构 * Department of Computer Engineering, Munich Institute of Robotics and Machine Intelligence (MIRMI), Technical University of Munich(计算机工程系,慕尼黑机器人与机器智能研究所(MIRMI),慕尼黑技术大学) Department of Aeronautics and Astronautics, Stanford University(航空与航天系,斯坦福大学)

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本文提出路径一致安全过滤方法,用于提升扩散策略在动态环境中的安全性和任务成功率。

Comments Accepted to IEEE ICRA 2026. Project page: https://tum-lsy.github.io/pacs/. 8 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01068 2026-03-11 cs.RO cs.LG 78%

Compose Your Policies! Improving Diffusion-based or Flow-based Robot Policies via Test-time Distribution-level Composition

制定你的策略!通过测试时的分布级组合改进扩散型或流型机器人策略

Jiahang Cao, Yize Huang, Hanzhong Guo, Rui Zhang, Mu Nan, Weijian Mai, Jiaxu Wang, Hao Cheng, Jingkai Sun, Gang Han, Wen Zhao, Qiang Zhang, Yijie Guo, Qihao Zheng, Chunfeng Song, Xiao Li, Ping Luo, Andrew F. Luo

机构 * The University of Hong Kong(香港大学) Beijing Innovation Center of Humanoid Robotics(北京人形机器人创新中心) Shanghai AI Lab(上海人工智能实验室) Shanghai Jiaotong University(上海交通大学) The Hong Kong University of Science and Technology(香港科技大学)

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本研究提出无需训练的通用策略组合方法,通过测试时分布级组合提升机器人策略性能。

Comments Accepted to ICLR 2026. Project Page: https://sagecao1125.github.io/GPC-Site/

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.10113 2026-03-11 math.PR 78%

Long-range one-dimensional internal diffusion-limited aggregation

长距离一维内部扩散限制聚集

Conrado da Costa, Debleena Thacker, Andrew Wade

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 该研究探讨了一维内部扩散限制聚集过程,在不同增量分布条件下分析簇结构的对称性和连续性,提出在有限方差和无限方差情况下的不同结论。

Comments 36 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09075 2026-03-11 eess.IV 78%

M2Diff: Multi-Modality Multi-Task Enhanced Diffusion Model for MRI-Guided Low-Dose PET Enhancement

M2Diff:多模态多任务增强扩散模型用于MRI引导的低剂量PET增强

Ghulam Nabi Ahmad Hassan Yar, Himashi Peiris, Victoria Mar, Cameron Dennis Pain, Zhaolin Chen

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 M2Diff通过多模态多任务扩散模型,利用MRI和低剂量PET扫描分别学习模态特定特征,并通过层次化特征融合重建标准剂量PET,从而提升重建保真度。

Comments Copyright 2026 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work in other works

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09063 2026-03-11 physics.app-ph physics.comp-ph 78%

A Stable, High-Order Time-Stepping Scheme for the Drift-Diffusion Model in Modern Solar Cell Simulation

一种用于现代太阳能电池模拟的稳定高阶时间推进方案

Jun Du, Jun Yan

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本文提出一种高阶时间推进方案,用于高精度模拟太阳能电池中的电荷、激子和离子传输。

Comments 16 pages, 12 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09039 2026-03-11 math.PR 78%

Critical stationary fluctuations in reaction--diffusion processes

反应-扩散过程中临界稳态波动

Luis Cardoso, Claudio Landim, Kenkichi Tsunoda

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 研究了一种反应-扩散过程在临界点的稳态波动,发现磁化率缩放后呈现非高斯波动,而密度场在更快模式上波动更小。

Comments 26 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.08949 2026-03-11 q-bio.NC 78%

Diffusion of Neuromodulators for Temporal Credit Assignment

神经调质的扩散用于时间信用分配

João Barretto-Bittar, Anna Levina, Emmanouil Giannakakis, Roxana Zeraati

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本研究提出了一种基于神经调质扩散的信用分配机制,通过局部误差扩散提升稀疏连接神经网络的学习能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.08901 2026-03-11 cs.CR cs.AI 78%

NetDiffuser: Deceiving DNN-Based Network Attack Detection Systems with Diffusion-Generated Adversarial Traffic

NetDiffuser: 通过扩散生成对抗性流量欺骗基于深度学习的网络攻击检测系统

Pratyay Kumar, Abu Saleh Md Tayeen, Satyajayant Misra, Huiping Cao, Jiefei Liu, Qixu Gong, Jayashree Harikumar

机构 * Department of Computer Science, New Mexico State University(新墨西哥州立大学计算机科学系) University of Hartford(哈特福大学) DEVCOM Analysis Center(DEVCOM分析中心)

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 NetDiffuser通过扩散模型生成自然对抗性示例,有效欺骗基于深度学习的网络攻击检测系统,提升攻击成功率并降低检测性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.08799 2026-03-11 quant-ph cs.NA math-ph math.MP math.NA 78%

Quantum algorithm for anisotropic diffusion and convection equations with vector norm scaling

量子算法用于各向异性扩散与对流方程的向量范数缩放

Julien Zylberman, Thibault Fredon, Nuno F. Loureiro, Fabrice Debbasch

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 该研究提出了一种量子算法,用于高效求解各向异性扩散与对流方程,通过向量范数分析实现了指数级的误差缩减。

Comments This preprint has not undergone peer review or any post-submission improvements or corrections. The Version of Record of this contribution is published in Quantum Engineering Sciences and Technologies for Industry and Services (QUEST-IS 2025), and is available online at https://doi.org/10.1007/978-3-032-13855-2_23

Journal ref International Conference on Quantum Engineering Sciences and Technologies for Industry and Services,255-263,2025,Springer

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.07357 2026-03-11 cs.LG cs.AI 67%

Latent Generative Models with Tunable Complexity for Compressed Sensing and other Inverse Problems

具有可调复杂度的潜在生成模型用于压缩感知及其他反问题

Sean Gunn, Jorio Cocola, Oliver De Candido, Vaggos Chatziafratis, Paul Hand

机构 * Northeastern University(东北大学) Harvard University(哈佛大学) Technical University of Munich(慕尼黑技术大学) UC Santa Cruz(加州大学圣克鲁兹分校)

专题命中 扩散模型 :diffusion(abstract);inpainting(abstract)

AI总结 本文提出可调复杂度生成模型,用于提升压缩感知等反问题的重建性能,并通过理论分析和实验验证其有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09042 2026-03-11 eess.IV 67%

Robust Wildfire Forecasting under Partial Observability: From Reconstruction to Prediction

在部分可观察性下实现鲁棒的野火预报:从重建到预测

Chen Yang, Mehdi Zafari, Ziheng Duan, A. Lee Swindlehurst

专题命中 扩散模型 :diffusion(abstract);inpainting(abstract)

AI总结 本文提出了一种两阶段概率框架,通过先恢复受损观测再预测野火动态,以解决部分可观察性下的野火预报问题,实验表明其在多种设置下均优于基线方法。

Comments 14 pages, 7 figures, and 4 tables. Submitted to IEEE for review. Codes and datasets available at: https://github.com/LS-Wireless/Robust-Wildfire-Forecasting

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09743 2026-03-11 cs.CV 57%

LAP: A Language-Aware Planning Model For Procedure Planning In Instructional Videos

LAP:一种语言感知的规划模型用于教学视频中的过程规划

Lei Shi, Victor Aregbede, Andreas Persson, Martin Längkvist, Amy Loutfi, Stephanie Lowry

机构 * Örebro University(奥雷布罗大学)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

AI总结 LAP通过利用语言描述提升过程规划的准确性,实现教学视频中动作序列的高效生成。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09702 2026-03-11 cs.CV 57%

TriFusion-SR: Joint Tri-Modal Medical Image Fusion and SR

TriFusion-SR:联合三模态医学图像融合与超分辨率

Fayaz Ali Dharejo, Sharif S. M. A., Aiman Khalil, Nachiket Chaudhary, Rizwan Ali Naqvi, Radu Timofte

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

AI总结 TriFusion-SR通过联合三模态医学图像融合与超分辨率技术,利用小波引导的条件扩散框架实现频率感知的跨模态交互,提升图像质量和分辨率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09236 2026-03-11 cs.CV cs.AI 57%

BridgeDiff: Bridging Human Observations and Flat-Garment Synthesis for Virtual Try-Off

BridgeDiff: 联接人体观察与平面服装合成以实现虚拟试穿

Shuang Liu, Ao Yu, Linkang Cheng, Xiwen Huang, Li Zhao, Junhui Liu, Zhiting Lin, Yu Liu

机构 * School of Integrated Circuits, Anhui University, Hefei, China(安徽大学集成电路学院,合肥,中国) Anhui Provincial High-performance Integrated Circuit Engineering Research Center(安徽省高性能集成电路工程研究中心) School of Astronautics, Northwestern Polytechnical University, Xi’an,China(西北工业大学航天学院,西安,中国)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

AI总结 BridgeDiff通过结合人体观察与平面服装合成,提升虚拟试穿中平面服装重建的质量与结构稳定性。

Comments 33 pages, 16 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.08483 2026-03-11 cs.CV cs.AI cs.LG 57%

X-AVDT: Audio-Visual Cross-Attention for Robust Deepfake Detection

X-AVDT:用于鲁棒深度伪造检测的音频视觉交叉注意力

Youngseo Kim, Kwan Yun, Seokhyeon Hong, Sihun Cha, Colette Suhjung Koo, Junyong Noh

机构 * Visual Media Lab, KAIST(韩国科学技术院视觉媒体实验室)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

AI总结 X-AVDT通过利用生成器内部的音频视觉一致性线索,提出了一种鲁棒且具有通用性的深度伪造检测方法,有效提升了检测性能。

Journal ref CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏