arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 69989 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 69989 篇

2412.10891 2024-12-18 cs.CV cs.LG 83%

Zigzag Diffusion Sampling: Diffusion Models Can Self-Improve via Self-Reflection

Lichen Bai, Shitong Shao, Zikai Zhou, Zipeng Qi, Zhiqiang Xu, Haoyi Xiong, Zeke Xie

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.08495 2024-12-18 cs.CV 83%

FunEditor: Achieving Complex Image Edits via Function Aggregation with Diffusion Models

Mohammadreza Samadi, Fred X. Han, Mohammad Salameh, Hao Wu, Fengyu Sun, Chunhua Zhou, Di Niu

专题命中 扩散模型 :diffusion(title,abstract);image editing(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.11710 2024-12-17 cs.CV cs.AI 83%

Re-Attentional Controllable Video Diffusion Editing

Yuanzhi Wang, Yong Li, Mengyi Liu, Xiaoya Zhang, Xin Liu, Zhen Cui, Antoni B. Chan

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments Accepted by AAAI 2025. Codes are released at: https://github.com/mdswyz/ReAtCo

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.11519 2024-12-17 cs.CV 83%

LineArt: A Knowledge-guided Training-free High-quality Appearance Transfer for Design Drawing with Diffusion Model

Xi Wang, Hongzhen Li, Heng Fang, Yichen Peng, Haoran Xie, Xi Yang, Chuntao Li

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Project Page: https://meaoxixi.github.io/LineArt/

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.11402 2024-12-17 cs.CV 83%

Video Diffusion Models are Strong Video Inpainter

Minhyeok Lee, Suhwan Cho, Chajin Shin, Jungho Lee, Sunghun Yang, Sangyoun Lee

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

Comments Accepted to AAAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.12868 2024-12-17 cs.CV 83%

DiffBoost: Enhancing Medical Image Segmentation via Text-Guided Diffusion Model

Zheyuan Zhang, Lanhong Yao, Bin Wang, Debesh Jha, Gorkem Durak, Elif Keles, Alpay Medetalibeyoglu, Ulas Bagci

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Comments Accepted by IEEE TRANSACTIONS ON MEDICAL IMAGING

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.10448 2024-12-17 cs.CV cs.AI 83%

Unlocking Visual Secrets: Inverting Features with Diffusion Priors for Image Reconstruction

Sai Qian Zhang, Ziyun Li, Chuan Guo, Saeed Mahloujifar, Deeksha Dangwal, Edward Suh, Barbara De Salvo, Chiao Liu

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.10122 2024-12-16 cs.CV 83%

The Art of Deception: Color Visual Illusions and Diffusion Models

Alex Gomez-Villa, Kai Wang, Alejandro C. Parraga, Bartlomiej Twardowski, Jesus Malo, Javier Vazquez-Corral, Joost van de Weijer

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.09618 2024-12-13 cs.CV 83%

EasyRef: Omni-Generalized Group Image Reference for Diffusion Models via Multimodal LLM

Zhuofan Zong, Dongzhi Jiang, Bingqi Ma, Guanglu Song, Hao Shao, Dazhong Shen, Yu Liu, Hongsheng Li

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Tech report

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.08635 2024-12-12 cs.CL cs.CV cs.LG 83%

Multimodal Latent Language Modeling with Next-Token Diffusion

Yutao Sun, Hangbo Bao, Wenhui Wang, Zhiliang Peng, Li Dong, Shaohan Huang, Jianyong Wang, Furu Wei

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.08480 2024-12-12 cs.CV cs.IR cs.LG 83%

InvDiff: Invariant Guidance for Bias Mitigation in Diffusion Models

Min Hou, Yueying Wu, Chang Xu, Yu-Hao Huang, Chenxi Bai, Le Wu, Jiang Bian

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments KDD 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.17462 2024-12-12 cs.CV cs.AI 83%

EvolvED: Evolutionary Embeddings to Understand the Generation Process of Diffusion Models

Vidya Prasad, Hans van Gorp, Christina Humer, Ruud J. G. van Sloun, Anna Vilanova, Nicola Pezzotti

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.01367 2024-12-11 cs.CV cs.LG 83%

Bigger is not Always Better: Scaling Properties of Latent Diffusion Models

Kangfu Mei, Zhengzhong Tu, Mauricio Delbracio, Hossein Talebi, Vishal M. Patel, Peyman Milanfar

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments Accepted to TMLR. Camera-ready version

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.16807 2024-12-11 cs.CV 83%

Negative-prompt Inversion: Fast Image Inversion for Editing with Text-guided Diffusion Models

Daiki Miyake, Akihiro Iohara, Yu Saito, Toshiyuki Tanaka

专题命中 扩散模型 :diffusion(title,abstract);image editing(abstract);分类 cs.CV

Comments 20 pages, 14 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.06578 2024-12-10 cs.CV 83%

MoViE: Mobile Diffusion for Video Editing

Adil Karjauv, Noor Fathima, Ioannis Lelekas, Fatih Porikli, Amir Ghodrati, Amirhossein Habibian

专题命中 扩散模型 :diffusion(title,abstract);image editing(abstract);分类 cs.CV

Comments 8 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.05984 2024-12-10 cs.CV 83%

Nested Diffusion Models Using Hierarchical Latent Priors

Xiao Zhang, Ruoxi Jiang, Rebecca Willett, Michael Maire

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.04364 2024-12-10 cs.CV cs.AI cs.LG 83%

VideoGuide: Improving Video Diffusion Models without Training Through a Teacher's Guide

Dohun Lee, Bryan S Kim, Geon Yeong Park, Jong Chul Ye

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments 26 pages, 19 figures, Project Page: https://dohunlee1.github.io/videoguide.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.08850 2024-12-10 cs.CV 83%

COVE: Unleashing the Diffusion Feature Correspondence for Consistent Video Editing

Jiangshan Wang, Yue Ma, Jiayi Guo, Yicheng Xiao, Gao Huang, Xiu Li

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments NeurIPS 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.03895 2024-12-06 cs.CV cs.AI cs.LG 83%

A Noise is Worth Diffusion Guidance

Donghoon Ahn, Jiwon Kang, Sanghyun Lee, Jaewon Min, Minjae Kim, Wooseok Jang, Hyoungwon Cho, Sayak Paul, SeonHwa Kim, Eunju Cha, Kyong Hwan Jin, Seungryong Kim

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Project page: https://cvlab-kaist.github.io/NoiseRefine/

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.03809 2024-12-06 cs.CV 83%

EditScout: Locating Forged Regions from Diffusion-based Edited Images with Multimodal LLM

Quang Nguyen, Truong Vu, Trong-Tung Nguyen, Yuxin Wen, Preston K Robinette, Taylor T Johnson, Tom Goldstein, Anh Tran, Khoi Nguyen

专题命中 扩散模型 :diffusion(title,abstract);image editing(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.09633 2024-12-06 cs.CV 83%

DuoDiff: Accelerating Diffusion Models with a Dual-Backbone Approach

Daniel Gallo Fernández, Răzvan-Andrei Matişan, Alejandro Monroy Muñoz, Ana-Maria Vasilcoiu, Janusz Partyka, Tin Hadži Veljković, Metod Jazbec

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Accepted to NeurIPS, see https://openreview.net/forum?id=G7E4tNmmHD

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.02432 2024-12-06 cs.CV 83%

Orthogonal Adaptation for Modular Customization of Diffusion Models

Ryan Po, Guandao Yang, Kfir Aberman, Gordon Wetzstein

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments Project page: https://ryanpo.com/ortha/ Hugging Face Demo: https://huggingface.co/spaces/ujin-song/ortha

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.07360 2024-12-05 cs.CV 83%

Boosting Latent Diffusion with Flow Matching

Johannes Schusterbauer, Ming Gui, Pingchuan Ma, Nick Stracke, Stefan A. Baumann, Vincent Tao Hu, Björn Ommer

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Comments ECCV 2024 (Oral), Project Page: https://compvis.github.io/fm-boosting/

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.02962 2024-12-05 cs.CV cs.DC 83%

Partially Conditioned Patch Parallelism for Accelerated Diffusion Model Inference

XiuYu Zhang, Zening Luo, Michelle E. Lu

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.01756 2024-12-05 cs.CV 83%

ImageFolder: Autoregressive Image Generation with Folded Tokens

Xiang Li, Kai Qiu, Hao Chen, Jason Kuen, Jiuxiang Gu, Bhiksha Raj, Zhe Lin

专题命中 扩散模型 :image generation(title,abstract);diffusion(abstract);分类 cs.CV

Comments Code: https://github.com/lxa9867/ImageFolder

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.02631 2024-12-04 cs.CV cs.LG 83%

Sharp-It: A Multi-view to Multi-view Diffusion Model for 3D Synthesis and Manipulation

Yiftach Edelstein, Or Patashnik, Dana Cohen-Bar, Lihi Zelnik-Manor

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments Project page at https://yiftachede.github.io/Sharp-It/

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.01244 2024-12-04 cs.CV 83%

Concept Replacer: Replacing Sensitive Concepts in Diffusion Models via Precision Localization

Lingyun Zhang, Yu Xie, Yanwei Fu, Ping Chen

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.01199 2024-12-03 cs.CV cs.AI cs.LG 83%

TinyFusion: Diffusion Transformers Learned Shallow

Gongfan Fang, Kunjun Li, Xinyin Ma, Xinchao Wang

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.01118 2024-12-03 cs.CV cs.CR 83%

LoyalDiffusion: A Diffusion Model Guarding Against Data Replication

Chenghao Li, Yuke Zhang, Dake Chen, Jingqi Xu, Peter A. Beerel

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments 13 pages, 6 figures, Submission to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.00665 2024-12-03 cs.CV 83%

Learning on Less: Constraining Pre-trained Model Learning for Generalizable Diffusion-Generated Image Detection

Yingjian Chen, Lei Zhang, Yakun Niu, Lei Tan, Pei Chen

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏