arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 69989 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 69989 篇

2505.19385 2025-11-26 cs.CV cs.AI 83%

Advancing Limited-Angle CT Reconstruction Through Diffusion-Based Sinogram Completion

通过扩散基于的sinogram补全推进有限角度CT重建

Jiaqi Guo, Santiago Lopez-Tapia, Aggelos K. Katsaggelos

机构 * Dept. of Electrical and Computer Engineering, Northwestern University, Evanston, IL, USA(电气与计算机工程系,西北大学,爱荷华州埃文斯顿)

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

AI总结 本文提出基于扩散模型的sinogram补全方法,结合蒸馏和伪逆约束,实现高效准确的有限角度CT重建,有效抑制伪影并保持结构细节。

Comments Accepted at the 2025 IEEE International Conference on Image Processing (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19434 2025-11-25 cs.CV cs.LG stat.ML 83%

Breaking the Likelihood-Quality Trade-off in Diffusion Models by Merging Pretrained Experts

通过合并预训练专家打破扩散模型中似然-质量权衡

Yasin Esfandiari, Stefan Bauer, Sebastian U. Stich, Andrea Dittadi

机构 * Saarland University(萨尔兰大学) Helmholtz AI(亥姆霍兹人工智能研究所) Technical University of Munich(慕尼黑技术大学) CISPA Helmholtz Center for Information Security(亥姆霍兹信息安全部分研究所) MPI for Intelligent Systems, Tübingen(图宾根智能系统研究所)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

AI总结 通过在去噪过程中切换预训练专家,该方法有效打破扩散模型中似然与质量的权衡,提升图像生成质量和似然

Comments ICLR 2025 DeLTa workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.01421 2025-11-25 cs.CV 83%

InfoScale: Unleashing Training-free Variable-scaled Image Generation via Effective Utilization of Information

InfoScale: 通过有效利用信息实现免训练可变尺度图像生成

Guohui Zhang, Jiangtong Tan, Linjiang Huang, Zhonghang Yuan, Mingde Yao, Jie Huang, Feng Zhao

机构 * USTC(中国科学技术大学) Beihang University(北京航空航天大学) CUHK MMLab(香港中文大学多媒体实验室) Kuaishou Technology(快手科技)

专题命中 扩散模型 :image generation(title,abstract);diffusion(abstract);分类 cs.CV

AI总结 InfoScale通过有效利用信息解决扩散模型在可变尺度图像生成中的信息丢失、聚合不灵活和分布不匹配问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.16904 2025-11-24 cs.CV 83%

Warm Diffusion: Recipe for Blur-Noise Mixture Diffusion Models

Warm Diffusion: Blur-Noise Mixture Diffusion Models的配方

Hao-Chien Hsueh, Chi-En Yen, Wen-Hsiao Peng, Ching-Chun Huang

机构 * Department of Computer Science, National Yang Ming Chiao Tung University, Hsinchu, Taiwan(计算机科学系,National Yang Ming Chiao Tung大学,Hsinchu,台湾)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

AI总结 Warm Diffusion通过融合模糊与噪声的混合模型,解决传统热扩散和冷扩散的不足,提升图像生成效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.10629 2025-11-24 cs.CV 83%

One Small Step in Latent, One Giant Leap for Pixels: Fast Latent Upscale Adapter for Your Diffusion Models

潜在空间中的一小步,像素世界中的一大跃:一种快速的潜在上采样适配器用于您的扩散模型

Aleksandr Razin, Danil Kazantsev, Ilya Makarov

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

AI总结 提出一种轻量级潜在上采样适配器,通过潜在空间单次前向传递实现高分辨率图像合成,提升效率并保持保真度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.16317 2025-11-21 cs.CV 83%

NaTex: Seamless Texture Generation as Latent Color Diffusion

NaTex: 无缝纹理生成作为潜在颜色扩散

Zeqiang Lai, Yunfei Zhao, Zibo Zhao, Xin Yang, Xin Huang, Jingwei Huang, Xiangyu Yue, Chunchao Guo

机构 * MMLab, CUHK(CUHK多媒体实验室) Tencent Hunyuan(腾讯文生)

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

AI总结 NaTex通过直接在3D空间中预测纹理颜色,提出潜在颜色扩散方法,实现高效且一致的纹理生成与重建。

Comments Technical Report

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.14113 2025-11-19 cs.CV 83%

Coffee: Controllable Diffusion Fine-tuning

Ziyao Zeng, Jingcheng Ni, Ruyi Liu, Alex Wong

机构 * Yale University(耶鲁大学) Brown University(布朗大学)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13689 2025-11-19 cs.CL cs.CV 83%

Crossing Borders: A Multimodal Challenge for Indian Poetry Translation and Image Generation

Sofia Jamil, Kotla Sai Charan, Sriparna Saha, Koustava Goswami, Joseph K J

专题命中 扩散模型 :image generation(title,abstract);diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.09611 2025-11-19 cs.CV 83%

MMaDA-Parallel: Multimodal Large Diffusion Language Models for Thinking-Aware Editing and Generation

Ye Tian, Ling Yang, Jiongfan Yang, Anran Wang, Yu Tian, Jiani Zheng, Haochen Wang, Zhiyang Teng, Zhuochen Wang, Yinjie Wang, Yunhai Tong, Mengdi Wang, Xiangtai Li

机构 * Peking University(北京大学) ByteDance(字节跳动) Princeton University(普林斯顿大学) CASIA(中国科学院自动化研究所) The University of Chicago(芝加哥大学)

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Comments Project Page: https://tyfeld.github.io/mmadaparellel.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.05772 2025-11-18 cs.CV 83%

MAISI-v2: Accelerated 3D High-Resolution Medical Image Synthesis with Rectified Flow and Region-specific Contrastive Loss

Can Zhao, Pengfei Guo, Dong Yang, Yucheng Tang, Yufan He, Benjamin Simon, Mason Belue, Stephanie Harmon, Baris Turkbey, Daguang Xu

专题命中 扩散模型 :image synthesis(title,abstract);diffusion(abstract);分类 cs.CV

Comments Accepted by AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.06157 2025-11-18 cs.CV 83%

3D-free meets 3D priors: Novel View Synthesis from a Single Image with Pretrained Diffusion Guidance

Taewon Kang, Divya Kothandaraman, Dinesh Manocha, Ming C. Lin

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Accepted to The 40th Annual AAAI Conference on Artificial Intelligence (AAAI-26), AAAI 2026 Workshop on AI for Environmental Science (AI4ES). Due to arXiv's 1,920-character limit, the abstract here is shortened. Please refer to the paper (View PDF) to read the full abstract. 14 pages, 13 figures, v5: AAAI-26 camera-ready version

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.11231 2025-11-17 cs.CV cs.AI 83%

3D Gaussian and Diffusion-Based Gaze Redirection

Abiram Panchalingam, Indu Bodala, Stuart Middleton

机构 * School of Electronics and Computer Science, University of Southampton(电子与计算机科学学院,南安普顿大学)

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.08207 2025-11-14 cs.CV cs.LG 83%

DICE: Discrete Inversion Enabling Controllable Editing for Multinomial Diffusion and Masked Generative Models

Xiaoxiao He, Quan Dao, Ligong Han, Song Wen, Minhao Bai, Di Liu, Han Zhang, Martin Renqiang Min, Felix Juefei-Xu, Chaowei Tan, Bo Liu, Kang Li, Hongdong Li, Junzhou Huang, Faez Ahmed, Akash Srivastava, Dimitris Metaxas

机构 * Rutgers University(新泽西罗格斯大学) MIT-IBM Watson AI Lab(MIT-IBM沃森人工智能实验室) Red Hat AI Innovation(红帽AI创新) Google DeepMind(谷歌DeepMind) NYU(纽约大学) Walmart Global Tech(沃尔玛全球技术) NEC Labs America(NEC美国实验室) Massachusetts Institute of Technology(麻省理工学院) ANU(澳大利亚国立大学) UT Arlington(德克萨斯大学阿灵顿分校)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Project webpage: https://hexiaoxiao-cs.github.io/DICE/. This paper was accepted to CVPR 2025 but later desk-rejected post camera-ready, due to a withdrawal from ICLR made 14 days before reviewer assignment

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08090 2025-11-12 cs.CV cs.AI 83%

StableMorph: High-Quality Face Morph Generation with Stable Diffusion

Wassim Kabbani, Kiran Raja, Raghavendra Ramachandra, Christoph Busch

机构 * Norwegian University of Science and Technology(挪威科学技术大学)

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Journal ref International Joint Conference on Biometrics 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07806 2025-11-12 cs.CV 83%

PC-Diffusion: Aligning Diffusion Models with Human Preferences via Preference Classifier

Shaomeng Wang, He Wang, Xiaolu Wei, Longquan Dai, Jinhui Tang

机构 * School of Computer Science and Engineering, Nanjing University of Science and Technology(计算机科学与工程学院,南京理工大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments 10 pages, 3 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07499 2025-11-12 cs.CV cs.AI 83%

Toward the Frontiers of Reliable Diffusion Sampling via Adversarial Sinkhorn Attention Guidance

Kwanyoung Kim

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments Accepted to AAAI 26

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.24434 2025-11-11 cs.LG cs.CV 83%

Graph Flow Matching: Enhancing Image Generation with Neighbor-Aware Flow Fields

Md Shahriar Rahim Siddiqui, Moshe Eliasof, Eldad Haber

专题命中 扩散模型 :image generation(title,abstract);diffusion(abstract);分类 cs.CV

Comments The 40th Annual AAAI Conference on Artificial Intelligence

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.18877 2025-11-11 cs.CV cs.CR cs.LG 83%

Mitigating Sexual Content Generation via Embedding Distortion in Text-conditioned Diffusion Models

Jaesin Ahn, Heechul Jung

机构 * Department of Artificial Intelligence(人工智能系) Kyungpook National University(全北国立大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments NeurIPS 2025 accepted. Official code: https://github.com/amoeba04/des

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.09935 2025-11-11 eess.IV cs.CV physics.med-ph 83%

Physics-informed DeepCT: Sinogram Wavelet Decomposition Meets Masked Diffusion

Zekun Zhou, Tan Liu, Bing Yu, Yanru Gong, Liu Shi, Qiegen Liu

机构 * School of Mathematics and Computer Sciences, Nanchang University, Nanchang, China(南昌大学数学与计算机科学学院) School of Information Engineering, Nanchang University, Nanchang, China(南昌大学信息工程学院) Key Laboratory of Advanced Medical Imaging and Intelligent Computing of Guizhou Province, China(贵州省先进医学影像与智能计算重点实验室)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04117 2025-11-07 cs.CV 83%

Tortoise and Hare Guidance: Accelerating Diffusion Model Inference with Multirate Integration

Yunghee Lee, Byeonghyun Pak, Junwha Hong, Hoseong Kim

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Comments 21 pages, 8 figures. NeurIPS 2025. Project page: https://yhlee-add.github.io/THG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27171 2025-11-06 cs.CV cs.AI 83%

H2-Cache: A Novel Hierarchical Dual-Stage Cache for High-Performance Acceleration of Generative Diffusion Models

Mingyu Sung, Il-Min Kim, Sangseok Yun, Jae-Mo Kang

机构 * Department of Artificial Intelligence, Kyungpook National University(人工智能系,庆尚国立大学) Department of Electrical and Computer Engineering, Queen’s University(电气与计算机工程系,皇后大学) Department of Information and Communications Engineering, Pukyong National University(信息与通信工程系,浦项国立大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.02462 2025-11-05 cs.CV 83%

KAO: Kernel-Adaptive Optimization in Diffusion for Satellite Image

Teerapong Panboonyuen

机构 * Chulalongkorn University(朱拉隆功大学)

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

Comments 18 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.15724 2025-11-05 cs.CV 83%

A Practical Investigation of Spatially-Controlled Image Generation with Transformers

Guoxuan Xia, Harleen Hanspal, Petru-Daniel Tudosiu, Shifeng Zhang, Sarah Parisot

机构 * Huawei Noah’s Ark Lab(华为诺亚实验室)

专题命中 扩散模型 :image generation(title,abstract);diffusion(abstract);分类 cs.CV

Comments TMLR https://openreview.net/forum?id=loT6xhgLYK

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01175 2025-11-05 cs.CV 83%

Diffusion Transformer meets Multi-level Wavelet Spectrum for Single Image Super-Resolution

Peng Du, Hui Li, Han Xu, Paul Barom Jeon, Dongwook Lee, Daehyun Ji, Ran Yang, Feng Zhu

机构 * Samsung R&D Institute China Xi’an (SRCX)(三星中国研发中心西安(SRCX)) Samsung Electronics Co., LTD., South Korea(三星电子有限公司,韩国)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments ICCV 2025 Oral Paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21015 2025-11-05 cs.CV cs.LG quant-ph 83%

MediQ-GAN: Quantum-Inspired GAN for High Resolution Medical Image Generation

Qingyue Jiao, Yongcan Tang, Jun Zhuang, Jason Cong, Yiyu Shi

机构 * Independent Researcher(独立研究者)

专题命中 扩散模型 :image generation(title,abstract);diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01645 2025-11-04 cs.CV 83%

Enhancing Diffusion-based Restoration Models via Difficulty-Adaptive Reinforcement Learning with IQA Reward

Xiaogang Xu, Ruihang Chu, Jian Wang, Kun Zhou, Wenjie Shu, Harry Yang, Ser-Nam Lim, Hao Chen, Liang Lin

机构 * The Chinese University of Hong Kong(香港中文大学) Tsinghua University(清华大学) Snap Research Shenzhen University(深圳大学) HKUST(香港科技大学) University of Central Florida(佛罗里达大学) UC Davis(加州大学戴维斯分校) Sun Yat-Sen University(孙中山大学)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01466 2025-11-04 cs.CV 83%

SecDiff: Diffusion-Aided Secure Deep Joint Source-Channel Coding Against Adversarial Attacks

Changyuan Zhao, Jiacheng Wang, Ruichen Zhang, Dusit Niyato, Hongyang Du, Zehui Xiong, Dong In Kim, Ping Zhang

机构 * College of Computing and Data Science, Nanyang Technological University(计算与数据科学学院,南洋理工大学) Department of Electrical and Electronic Engineering, University of Hong Kong(电子与电气工程系,香港大学) School of Electronics, Electrical Engineering and Computer Science, Queen’s University Belfast(电子、电气工程与计算机科学学院,女王学院贝尔法斯特分校) Department of Electrical and Computer Engineering, Sungkyunkwan University(电气与计算机工程系,成均馆大学) State Key Laboratory of Networking and Switching Technology, Beijing University of Posts and Telecommunications(网络与交换技术国家重点实验室,北京邮电大学)

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

Comments 13 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.07140 2025-11-04 cs.CV 83%

FIRE: Robust Detection of Diffusion-Generated Images via Frequency-Guided Reconstruction Error

Beilin Chu, Xuan Xu, Xin Wang, Yufei Zhang, Weike You, Linna Zhou

机构 * School of CyberSpace Security, Beijing University of Posts and Telecommunications(网络安全学院,北京邮电大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments 14 pages, 14 figures. Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00429 2025-11-04 cs.CV cs.AI 83%

Enhancing Frequency Forgery Clues for Diffusion-Generated Image Detection

Daichi Zhang, Tong Zhang, Shiming Ge, Sabine Süsstrunk

机构 * School of Computer and Communication Sciences, EPFL(瑞士联邦理工学院计算机与通信科学学院) Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所) University of Chinese Academy of Sciences(中国科学院大学)

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.10978 2025-11-04 cs.CV cs.AI cs.LG 83%

Where and How to Perturb: On the Design of Perturbation Guidance in Diffusion and Flow Models

Donghoon Ahn, Jiwon Kang, Sanghyun Lee, Minjae Kim, Jaewon Min, Wooseok Jang, Sangwu Lee, Sayak Paul, Susung Hong, Seungryong Kim

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments Accepted at NeurIPS 2025. Project page: https://cvlab-kaist.github.io/HeadHunter/

详情

展开后加载摘要…

URL PDF HTML 收藏