arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 69989 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 69989 篇

2509.06573 2025-09-09 cs.GR 83%

From Rigging to Waving: 3D-Guided Diffusion for Natural Animation of Hand-Drawn Characters

Jie Zhou, Linzi Qu, Miu-Ling Lam, Hongbo Fu

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.GR

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.03221 2025-09-09 cs.CV eess.IV 83%

ADIR: Adaptive Diffusion for Image Reconstruction

Shady Abu-Hussein, Tom Tirer, Raja Giryes

机构 * Tel Aviv University(特拉维夫大学) Bar Ilan University(巴伊兰大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Project page https://shadyabh.github.io/ADIR/

Journal ref BMVC 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.16507 2025-09-05 cs.CV cs.LG 83%

Straighter Flow Matching via a Diffusion-Based Coupling Prior

Siyu Xing, Jie Cao, Huaibo Huang, Haichao Shi, Xiao-Yu Zhang

机构 * Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所) School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络空间安全学院) Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.02161 2025-09-03 cs.CV 83%

Enhancing Zero-Shot Pedestrian Attribute Recognition with Synthetic Data Generation: A Comparative Study with Image-To-Image Diffusion Models

Pablo Ayuso-Albizu, Juan C. SanMiguel, Pablo Carballeira

机构 * Universidad Autónoma de Madrid(西班牙马德里自治大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Paper accepted at AVSS 2025 conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00378 2025-09-03 cs.CV 83%

NoiseCutMix: A Novel Data Augmentation Approach by Mixing Estimated Noise in Diffusion Models

Shumpei Takezaki, Ryoma Bise, Shinnosuke Matsuo

机构 * Kyushu University(九州大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Accepted at ICCV2025 Workshop LIMIT

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20182 2025-08-29 cs.CV 83%

SDiFL: Stable Diffusion-Driven Framework for Image Forgery Localization

Yang Su, Shunquan Tan, Jiwu Huang

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.14209 2025-08-29 cs.LG cs.CV 83%

Unlearning Concepts from Text-to-Video Diffusion Models

Shiqi Liu, Yihua Tan

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20020 2025-08-28 cs.CV 83%

GS: Generative Segmentation via Label Diffusion

Yuhao Chen, Shubin Chen, Liang Lin, Guangrun Wang

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments 12 pages, 7 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18235 2025-08-26 cs.CV 83%

Sealing The Backdoor: Unlearning Adversarial Text Triggers In Diffusion Models Using Knowledge Distillation

Ashwath Vaithinathan Aravindan, Abha Jha, Matthew Salaway, Atharva Sandeep Bhide, Duygu Nur Yaldiz

机构 * University of Southern California(南加州大学)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17844 2025-08-26 cs.CV cs.LG 83%

Diffusion-Based Data Augmentation for Medical Image Segmentation

Maham Nazir, Muhammad Aqeel, Francesco Setti

机构 * School of Computer Science and Engineering, Beihang University(北京航空航天大学计算机科学与工程学院) Dept. of Engineering for Innovation Medicine, University of Verona(威尼斯大学创新医学工程系)

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

Comments Accepted to CVAMD Workshop at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17614 2025-08-26 cs.CV 83%

JCo-MVTON: Jointly Controllable Multi-Modal Diffusion Transformer for Mask-Free Virtual Try-on

Aowen Wang, Wei Li, Hao Luo, Mengxing Ao, Chenyu Zhu, Xinyang Li, Fan Wang

机构 * DAMO Academy, Alibaba Group(达摩院,阿里巴巴集团) Hupan Lab(汇安实验室) Zhejiang University(浙江大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17045 2025-08-26 cs.CV 83%

Styleclone: Face Stylization with Diffusion Based Data Augmentation

Neeraj Matiyali, Siddharth Srivastava, Gaurav Sharma

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16212 2025-08-26 cs.CV cs.AI cs.LG 83%

OmniCache: A Trajectory-Oriented Global Perspective on Training-Free Cache Reuse for Diffusion Transformer Models

Huanpeng Chu, Wei Wu, Guanyu Fen, Yutao Zhang

机构 * Zhipu AI(智谱AI)

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.11688 2025-08-26 cs.CR cs.AI cs.MM 83%

Watermarking Visual Concepts for Diffusion Models

Liangqi Lei, Keke Gai, Jing Yu, Liehuang Zhu, Qi Wu

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16211 2025-08-25 cs.CV 83%

Forecast then Calibrate: Feature Caching as ODE for Efficient Diffusion Transformers

Shikang Zheng, Liang Feng, Xinyu Wang, Qinming Zhou, Peiliang Cai, Chang Zou, Jiacheng Liu, Yuqi Lin, Junjie Chen, Yue Ma, Linfeng Zhang

机构 * Shanghai Jiao Tong University(上海交通大学) South China University of Technology(华南理工大学) Fudan University(复旦大学) Tsinghua University(清华大学) Hong Kong University of Science and Technology(香港科技大学)

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16158 2025-08-25 cs.CV 83%

RAGSR: Regional Attention Guided Diffusion for Image Super-Resolution

Haodong He, Yancheng Bai, Rui Lan, Xu Duan, Lei Sun, Xiangxiang Chu, Gui-Song Xia

机构 * School of Computer Science, Wuhan University(武汉大学计算机学院) Amap, Alibaba Group(阿里巴巴集团阿里的)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.08927 2025-08-21 cs.CV 83%

Dynamic watermarks in images generated by diffusion models

Yunzhuo Chen, Naveed Akhtar, Nur Al Hasan Haldar, Ajmal Mian

机构 * The University of Western Australia(西澳大学) The University of Melbourne(墨尔本大学) Curtin University(Curtin 大学)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.09482 2025-08-21 cs.CV 83%

Marrying Autoregressive Transformer and Diffusion with Multi-Reference Autoregression

Dingcheng Zhen, Qian Qiao, Xu Zheng, Tan Yu, Kangxi Wu, Ziwei Zhang, Siyuan Liu, Shunshun Yin, Ming Tao

机构 * Soul AI ICT Chinese Academy of Sciences(中国科学院)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.13737 2025-08-21 cs.CV 83%

RNDiff: Rainfall nowcasting with Condition Diffusion Model

Xudong Ling, Chaorong Li, Fengqing Qin, Peng Yang, Yuanyuan Huang

机构 * Faculty of Artificial Intelligence and Big Data(人工智能与大数据学院) Chongqing University of Technology(重庆理工大学) Yibin University(宜春大学) Chengdu University of Information Technology(成都信息工程大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.11550 2025-08-18 cs.CV 83%

Training-Free Anomaly Generation via Dual-Attention Enhancement in Diffusion Model

Zuo Zuo, Jiahao Dong, Yanyun Qu, Zongze Wu

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.13219 2025-08-15 cs.CV 83%

PiT: Progressive Diffusion Transformer

Jiafu Wu, Yabiao Wang, Jian Li, Jinlong Peng, Yun Cao, Chengjie Wang, Jiangning Zhang

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09968 2025-08-14 cs.LG cs.CV 83%

Noise Hypernetworks: Amortizing Test-Time Compute in Diffusion Models

Luca Eyring, Shyamgopal Karthik, Alexey Dosovitskiy, Nataniel Ruiz, Zeynep Akata

机构 * Technical University of Munich(慕尼黑技术大学) Munich Center of Machine Learning(慕尼黑机器学习中心) Helmholtz Munich(海德堡-慕尼黑 Helmholtz 中心) University of Tübingen(图宾根大学) Inceptive(Inceptive 公司) Google(谷歌公司)

专题命中 扩散模型 :diffusion(title,abstract);generative vision(abstract);分类 cs.CV

Comments Project page: https://noisehypernetworks.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07747 2025-08-12 cs.CV 83%

Grouped Speculative Decoding for Autoregressive Image Generation

Junhyuk So, Juncheol Shin, Hyunho Kook, Eunhyeok Park

机构 * Department of Computer Science and Engineering, POSTECH(计算机科学与工程系,POSTECH) Graduate School of Artificial Intelligence, POSTECH(人工智能研究生院,POSTECH)

专题命中 扩散模型 :image generation(title,abstract);diffusion(abstract);分类 cs.CV

Comments Accepted to the ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07183 2025-08-12 cs.HC cs.AI cs.LG cs.MM 83%

Explainability-in-Action: Enabling Expressive Manipulation and Tacit Understanding by Bending Diffusion Models in ComfyUI

Ahmed M. Abuzuraiq, Philippe Pasquier

机构 * School of Interactive Arts and Technology(交互艺术与技术学院) Simon Fraser University(西蒙弗雷泽大学)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.MM

Comments In Proceedings of Explainable AI for the Arts Workshop 2025 (XAIxArts 2025) arXiv:2406.14485

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.05954 2025-08-11 cs.CV cs.AI cs.CL 83%

Bifrost-1: Bridging Multimodal LLMs and Diffusion Models with Patch-level CLIP Latents

Han Lin, Jaemin Cho, Amir Zadeh, Chuan Li, Mohit Bansal

机构 * UNC Chapel Hill(北卡罗来纳大学教堂山分校) Lambda

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Project Page: https://bifrost-1.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.05755 2025-08-11 cs.CV cs.AI 83%

UnGuide: Learning to Forget with LoRA-Guided Diffusion Models

Agnieszka Polowczyk, Alicja Polowczyk, Dawid Malarz, Artur Kasymov, Marcin Mazur, Jacek Tabor, Przemysław Spurek

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15877 2025-08-08 cs.CV 83%

Repurposing 2D Diffusion Models with Gaussian Atlas for 3D Generation

Tiange Xiang, Kai Li, Chengjiang Long, Christian Häne, Peihong Guo, Scott Delp, Ehsan Adeli, Li Fei-Fei

机构 * Stanford University(斯坦福大学) Meta Reality Labs(Meta现实实验室)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.14404 2025-08-06 cs.CV cs.AI 83%

Causally Steered Diffusion for Automated Video Counterfactual Generation

Nikos Spyrou, Athanasios Vlontzos, Paraskevas Pegios, Thomas Melistas, Nefeli Gkouti, Yannis Panagakis, Giorgos Papanastasiou, Sotirios A. Tsaftaris

机构 * Spotify, UK(英国Spotify)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.00482 2025-08-06 cs.CV cs.AI 83%

JointDiT: Enhancing RGB-Depth Joint Modeling with Diffusion Transformers

Kwon Byung-Ki, Qi Dai, Lee Hyoseok, Chong Luo, Tae-Hyun Oh

机构 * POSTECH Microsoft Research Asia(微软亚洲研究院) KAIST(韩国科学技术院)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Accepted to IEEE/CVF International Conference on Computer Vision (ICCV) 2025. Project page: https://byungki-k.github.io/JointDiT/ Code: https://github.com/kaist-ami/JointDiT

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.14432 2025-08-06 cs.CV eess.IV 83%

IntroStyle: Training-Free Introspective Style Attribution using Diffusion Features

Anand Kumar, Jiteng Mu, Nuno Vasconcelos

机构 * University of California, San Diego(加州大学圣地亚哥分校)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments 17 pages, 16 figures

Journal ref Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏