arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 69989 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 69989 篇

2503.23538 2025-04-01 cs.CV 83%

Enhancing Creative Generation on Stable Diffusion-based Models

Jiyeon Han, Dahee Kwon, Gayoung Lee, Junho Kim, Jaesik Choi

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments CVPR 2025 accepted paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06930 2025-04-01 cs.CV 83%

Post-Training Quantization for Diffusion Transformer via Hierarchical Timestep Grouping

Ning Ding, Jing Han, Yuchuan Tian, Chao Xu, Kai Han, Yehui Tang

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.16778 2025-04-01 cs.CV 83%

RoomPainter: View-Integrated Diffusion for Consistent Indoor Scene Texturing

Zhipeng Huang, Wangbo Yu, Xinhua Cheng, ChengShu Zhao, Yunyang Ge, Mingyi Guo, Li Yuan, Yonghong Tian

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.11971 2025-04-01 cs.LG cs.AI cs.CV 83%

DDIL: Diversity Enhancing Diffusion Distillation With Imitation Learning

Risheek Garrepalli, Shweta Mahajan, Munawar Hayat, Fatih Porikli

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.22496 2025-03-31 cs.RO cs.CV 83%

Scenario Dreamer: Vectorized Latent Diffusion for Generating Driving Simulation Environments

Luke Rowe, Roger Girgis, Anthony Gosselin, Liam Paull, Christopher Pal, Felix Heide

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.03006 2025-03-28 cs.CV 83%

Frequency-Controlled Diffusion Model for Versatile Text-Guided Image-to-Image Translation

Xiang Gao, Zhengbo Xu, Junhan Zhao, Jiaying Liu

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments Proceedings of the 38th AAAI Conference on Artificial Intelligence (AAAI 2024)

Journal ref Proceedings of the AAAI Conference on Artificial Intelligence, 2024, 38(3), 1824-1832

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20484 2025-03-27 cs.CV cs.AI 83%

Contrastive Learning Guided Latent Diffusion Model for Image-to-Image Translation

Qi Si, Bo Wang, Zhao Zhang

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments 11 pages, 13 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19791 2025-03-26 cs.CV 83%

SITA: Structurally Imperceptible and Transferable Adversarial Attacks for Stylized Image Generation

Jingdan Kang, Haoxin Yang, Yan Cai, Huaidong Zhang, Xuemiao Xu, Yong Du, Shengfeng He

专题命中 扩散模型 :image generation(title,abstract);diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19731 2025-03-26 cs.CV 83%

PCM : Picard Consistency Model for Fast Parallel Sampling of Diffusion Models

Junhyuk So, Jiwoong Shin, Chaeyeon Jang, Eunhyeok Park

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Accepted to the CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19012 2025-03-26 cs.CV 83%

DiffV2IR: Visible-to-Infrared Diffusion Model via Vision-Language Understanding

Lingyan Ran, Lidong Wang, Guangcong Wang, Peng Wang, Yanning Zhang

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Comments Project page: https://diffv2ir.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.00759 2025-03-26 cs.CV 83%

DyMO: Training-Free Diffusion Model Alignment with Dynamic Multi-Objective Scheduling

Xin Xie, Dong Gong

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17074 2025-03-25 cs.CV 83%

Zero-Shot Styled Text Image Generation, but Make It Autoregressive

Vittorio Pippi, Fabio Quattrini, Silvia Cascianelli, Alessio Tonioni, Rita Cucchiara

专题命中 扩散模型 :image generation(title,abstract);diffusion(abstract);分类 cs.CV

Comments Accepted at CVPR2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18134 2025-03-25 cs.CV 83%

An Image-like Diffusion Method for Human-Object Interaction Detection

Xiaofei Hui, Haoxuan Qu, Hossein Rahmani, Jun Liu

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.13401 2025-03-25 cs.CV 83%

Zero-Shot Low Light Image Enhancement with Diffusion Prior

Joshua Cho, Sara Aghajanzadeh, Zhen Zhu, D. A. Forsyth

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.19324 2025-03-25 cs.CV cs.LG stat.ML 83%

Simpler Diffusion (SiD2): 1.5 FID on ImageNet512 with pixel-space diffusion

Emiel Hoogeboom, Thomas Mensink, Jonathan Heek, Kay Lamerigts, Ruiqi Gao, Tim Salimans

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.05399 2025-03-25 cs.CV cs.LG 83%

Sequential Posterior Sampling with Diffusion Models

Tristan S. W. Stevens, Oisín Nolan, Jean-Luc Robert, Ruud J. G. van Sloun

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Comments 5 pages, 4 figures, preprint

Journal ref 2025 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.00684 2025-03-25 cs.CV cs.CL 83%

Deciphering Oracle Bone Language with Diffusion Models

Haisu Guan, Huanxin Yang, Xinyu Wang, Shengwei Han, Yongge Liu, Lianwen Jin, Xiang Bai, Yuliang Liu

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments ACL 2024 Best Paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.01274 2025-03-25 cs.CV 83%

Diffusion Models with Deterministic Normalizing Flow Priors

Mohsen Zand, Ali Etemad, Michael Greenspan

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments 17 pages, 7 figures, Published in Transactions on Machine Learning Research (TMLR)

Journal ref https://openreview.net/pdf?id=ACMNVwcR6v, Transactions on Machine Learning Research (TMLR), 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17155 2025-03-24 cs.CV 83%

D2C: Unlocking the Potential of Continuous Autoregressive Image Generation with Discrete Tokens

Panpan Wang, Liqiang Niu, Fandong Meng, Jinan Xu, Yufeng Chen, Jie Zhou

专题命中 扩散模型 :image generation(title,abstract);diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.16921 2025-03-24 cs.CV cs.AI 83%

When Preferences Diverge: Aligning Diffusion Models with Minority-Aware Adaptive DPO

Lingfan Zhang, Chen Liu, Chengming Xu, Kai Hu, Donghao Luo, Chengjie Wang, Yanwei Fu, Yuan Yao

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.07001 2025-03-21 cs.CV cs.AI cs.LG 83%

From Image to Video: An Empirical Study of Diffusion Representations

Pedro Vélez, Luisa F. Polanía, Yi Yang, Chuhan Zhang, Rishabh Kabra, Anurag Arnab, Mehdi S. M. Sajjadi

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.14358 2025-03-19 cs.CV cs.LG 83%

RFMI: Estimating Mutual Information on Rectified Flow for Text-to-Image Alignment

Chao Wang, Giulio Franzese, Alessandro Finamore, Pietro Michiardi

专题命中 扩散模型 :text-to-image(title,abstract);diffusion(abstract);分类 cs.CV

Comments to appear at ICLR 2025 Workshop on Deep Generative Model in Machine Learning: Theory, Principle and Efficacy

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.13945 2025-03-19 cs.CV 83%

Make the Most of Everything: Further Considerations on Disrupting Diffusion-based Customization

Long Tang, Dengpan Ye, Sirun Chen, Xiuwen Shi, Yunna Lv, Ziyi Liu

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.13661 2025-03-19 cs.SE cs.AI cs.CV cs.LG 83%

Efficient Domain Augmentation for Autonomous Driving Testing Using Diffusion Models

Luciano Baresi, Davide Yi Xian Hu, Andrea Stocco, Paolo Tonella

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

Comments Accepted for publication at the 47th International Conference on Software Engineering (ICSE 2025). This research was partially supported by project EMELIOT, funded by MUR under the PRIN 2020 program (n. 2020W3A5FY), by the Bavarian Ministry of Economic Affairs, Regional Development and Energy, by the TUM Global Incentive Fund, and by the EU Project Sec4AI4Sec (n. 101120393)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.15304 2025-03-19 cs.LG cs.CV 83%

Unlearning Concepts in Diffusion Model via Concept Domain Correction and Concept Preserving Gradient

Yongliang Wu, Shiji Zhou, Mingzhuo Yang, Lianzhe Wang, Heng Chang, Wenbo Zhu, Xinting Hu, Xiao Zhou, Xu Yang

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments AAAI 2025 camera-ready version

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.12472 2025-03-18 cs.CV 83%

Diffusion-based Synthetic Data Generation for Visible-Infrared Person Re-Identification

Wenbo Dai, Lijing Lu, Zhihang Li

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Comments AAAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.01794 2025-03-18 cs.CV cs.AI 83%

IQA-Adapter: Exploring Knowledge Transfer from Image Quality Assessment to Diffusion-based Generative Models

Khaled Abud, Sergey Lavrushkin, Alexey Kirillov, Dmitriy Vatolin

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments GitHub repo: https://github.com/X1716/IQA-Adapter

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.10687 2025-03-17 cs.CV 83%

Context-guided Responsible Data Augmentation with Diffusion Models

Khawar Islam, Naveed Akhtar

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments ICLRw

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.07026 2025-03-17 cs.CV cs.AI 83%

Erase Diffusion: Empowering Object Removal Through Calibrating Diffusion Pathways

Yi Liu, Hao Zhou, Wenxiang Shang, Ran Lin, Benlei Cui

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

Comments accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.06664 2025-03-17 cs.CV cs.AI 83%

Decouple-Then-Merge: Finetune Diffusion Models as Multi-Task Learning

Qianli Ma, Xuefei Ning, Dongrui Liu, Li Niu, Linfeng Zhang

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Accepted by CVPR2025

详情

展开后加载摘要…

URL PDF HTML 收藏