arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 70015 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 70015 篇

2403.11415 2024-09-24 cs.CV cs.AI cs.LG 83%

DreamSampler: Unifying Diffusion Sampling and Score Distillation for Image Manipulation

Jeongsol Kim, Geon Yeong Park, Jong Chul Ye

专题命中 扩散模型 :diffusion(title,abstract);image editing(abstract);分类 cs.CV

Comments ECCV 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.12539 2024-09-20 cs.CV 83%

Improving Cone-Beam CT Image Quality with Knowledge Distillation-Enhanced Diffusion Model in Imbalanced Data Settings

Joonil Hwang, Sangjoon Park, NaHyeon Park, Seungryong Cho, Jin Sung Kim

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments MICCAI 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.19140 2024-09-19 cs.CV cs.AI 83%

QNCD: Quantization Noise Correction for Diffusion Models

Huanpeng Chu, Wei Wu, Chengjie Zang, Kun Yuan

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Comments Accepted by ACMMM2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.11831 2024-09-19 cs.RO cs.CV cs.LG 83%

RaggeDi: Diffusion-based State Estimation of Disordered Rags, Sheets, Towels and Blankets

Jikai Ye, Wanze Li, Shiraz Khan, Gregory S. Chirikjian

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.11689 2024-09-19 cs.CV cs.AI 83%

GUNet: A Graph Convolutional Network United Diffusion Model for Stable and Diversity Pose Generation

Shuowen Liang, Sisi Li, Qingyun Wang, Cen Zhang, Kaiquan Zhu, Tian Yang

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.09569 2024-09-17 cs.LG cs.CV cs.CY 83%

Bias Begets Bias: The Impact of Biased Embeddings on Diffusion Models

Sahil Kuchlous, Marvin Li, Jeffrey G. Wang

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments 19 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.09144 2024-09-17 cs.CV 83%

PrimeDepth: Efficient Monocular Depth Estimation with a Stable Diffusion Preimage

Denis Zavadski, Damjan Kalšan, Carsten Rother

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.10647 2024-09-17 cs.CV cs.AI cs.LG 83%

A Survey on Video Diffusion Models

Zhen Xing, Qijun Feng, Haoran Chen, Qi Dai, Han Hu, Hang Xu, Zuxuan Wu, Yu-Gang Jiang

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.08278 2024-09-13 cs.CV 83%

DreamHOI: Subject-Driven Generation of 3D Human-Object Interactions with Diffusion Priors

Thomas Hanwen Zhu, Ruining Li, Tomas Jakab

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments Project page: https://DreamHOI.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.15288 2024-09-13 cs.CV cs.LG 83%

Memory-Efficient 3D Denoising Diffusion Models for Medical Image Processing

Florentin Bieder, Julia Wolleb, Alicia Durrer, Robin Sandkühler, Philippe C. Cattin

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Accepted at MIDL 2023

Journal ref Medical Imaging with Deep Learning, PMLR 227:552-567, 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.07269 2024-09-12 cs.CV 83%

Realistic and Efficient Face Swapping: A Unified Approach with Diffusion Models

Sanoojan Baliah, Qinliang Lin, Shengcai Liao, Xiaodan Liang, Muhammad Haris Khan

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

Comments Accepted as a conference paper at WACV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.05414 2024-09-10 cs.CR cs.AI cs.CV 83%

CipherDM: Secure Three-Party Inference for Diffusion Model Sampling

Xin Zhao, Xiaojun Chen, Xudong Chen, He Li, Tingyu Fan, Zhendong Zhao

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.11633 2024-09-10 cs.CV 83%

Scaling Diffusion Transformers to 16 Billion Parameters

Zhengcong Fei, Mingyuan Fan, Changqian Yu, Debang Li, Junshi Huang

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.12908 2024-09-10 cs.CV cs.LG eess.IV 83%

Robust CLIP-Based Detector for Exposing Diffusion Model-Generated Images

Santosh, Li Lin, Irene Amerini, Xin Wang, Shu Hu

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.03198 2024-09-06 cs.CV 83%

RoomDiffusion: A Specialized Diffusion Model in the Interior Design Industry

Zhaowei Wang, Ying Hao, Hao Wei, Qing Xiao, Lulu Chen, Yulong Li, Yue Yang, Tianyi Li

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.01014 2024-09-04 cs.CV cs.AI 83%

From Bird's-Eye to Street View: Crafting Diverse and Condition-Aligned Images with Latent Diffusion Model

Xiaojie Xu, Tianshuo Xu, Fulong Ma, Yingcong Chen

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Accepted at International Conference on Robotics and Automation(ICRA)

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.08822 2024-09-04 cs.CV 83%

Planning and Rendering: Towards Product Poster Generation with Diffusion Models

Zhaochen Li, Fengheng Li, Wei Feng, Honghe Zhu, Yaoyu Li, Zheng Zhang, Jingjing Lv, Junjie Shen, Zhangang Lin, Jingping Shao, Zhenglu Yang

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.03517 2024-09-04 cs.CV cs.AI 83%

FRDiff : Feature Reuse for Universal Training-free Acceleration of Diffusion Models

Junhyuk So, Jungwon Lee, Eunhyeok Park

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Accepted at ECCV 2024. Code : https://github.com/ECoLab-POSTECH/FRDiff

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.01464 2024-09-02 physics.med-ph cs.CV eess.IV physics.comp-ph 83%

CT Reconstruction using Diffusion Posterior Sampling conditioned on a Nonlinear Measurement Model

Shudong Li, Xiao Jiang, Matthew Tivnan, Grace J. Gang, Yuan Shen, J. Webster Stayman

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments 24 pages, 12 figures, 1 table, submitted to SPIE Journal of Medical Imaging. Updated with more realistic phantom data, Poisson likelihood, and additional evaluations including hallucination evaluation, performance under multiple noise levels, inference time evaluation, and etc. Changes in authorship is based on unanimous agreement to acknowledge the adding authors' contributions in this work

Journal ref Journal of Medical Imaging 11(4), 043504 (2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.16467 2024-08-30 cs.NE cs.CV 83%

Spiking Diffusion Models

Jiahang Cao, Hanzhong Guo, Ziqing Wang, Deming Zhou, Hao Cheng, Qiang Zhang, Renjing Xu

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Accepted by IEEE Transactions on Artificial Intelligence

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.15890 2024-08-29 cs.CV 83%

Disentangled Diffusion Autoencoder for Harmonization of Multi-site Neuroimaging Data

Ayodeji Ijishakin, Ana Lawry Aguila, Elizabeth Levitis, Ahmed Abdulaal, Andre Altmann, James Cole

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.14780 2024-08-29 cs.CV 83%

Customize-A-Video: One-Shot Motion Customization of Text-to-Video Diffusion Models

Yixuan Ren, Yang Zhou, Jimei Yang, Jing Shi, Difan Liu, Feng Liu, Mingi Kwon, Abhinav Shrivastava

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments Accepted by ECCV 2024. Project page: https://customize-a-video.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.09091 2024-08-29 cs.CV 83%

Edit Temporal-Consistent Videos with Image Diffusion Model

Yuanzhi Wang, Yong Li, Xiaoya Zhang, Xin Liu, Anbo Dai, Antoni B. Chan, Zhen Cui

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments 10 pages, 7 figures

Journal ref ACM TOMM 2024, Codes: https://github.com/mdswyz/TCVE

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.13858 2024-08-27 cs.CV cs.LG 83%

Draw Like an Artist: Complex Scene Generation with Diffusion Model via Composition, Painting, and Retouching

Minghao Liu, Le Zhang, Yingjie Tian, Xiaochao Qu, Luoqi Liu, Ting Liu

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.20171 2024-08-27 cs.CV 83%

Diffusion Feedback Helps CLIP See Better

Wenxuan Wang, Quan Sun, Fan Zhang, Yepeng Tang, Jing Liu, Xinlong Wang

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.11326 2024-08-25 cs.LG cs.CV 83%

On the Trajectory Regularity of ODE-based Diffusion Sampling

Defang Chen, Zhenyu Zhou, Can Wang, Chunhua Shen, Siwei Lyu

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments ICML 2024, 30 pages. arXiv admin note: text overlap with arXiv:2305.19947

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.09731 2024-08-22 eess.IV cs.CV 83%

Reconstruct Spine CT from Biplanar X-Rays via Diffusion Learning

Zhi Qiao, Xuhui Liu, Xiaopeng Wang, Runkun Liu, Xiantong Zhen, Pei Dong, Zhen Qian

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.07969 2024-08-20 cs.CV eess.IV 83%

Fast Inference in Denoising Diffusion Models via MMD Finetuning

Emanuele Aiello, Diego Valsesia, Enrico Magli

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.08184 2024-08-16 cs.CV cs.LG 83%

Not Every Image is Worth a Thousand Words: Quantifying Originality in Stable Diffusion

Adi Haviv, Shahar Sarfaty, Uri Hacohen, Niva Elkin-Koren, Roi Livni, Amit H Bermano

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments GenLaw ICML 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.07516 2024-08-16 cs.CV eess.IV 83%

DIffSteISR: Harnessing Diffusion Prior for Superior Real-world Stereo Image Super-Resolution

Yuanbo Zhou, Xinlin Zhang, Wei Deng, Tao Wang, Tao Tan, Qinquan Gao, Tong Tong

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏