arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 4201 信号源:cs.CV, cs.GR, cs.MM

1. 可控生成 4201 篇

2412.01792 2024-12-03 cs.CV cs.GR 81%

CTRL-D: Controllable Dynamic 3D Scene Editing with Personalized 2D Diffusion

Kai He, Chin-Hsuan Wu, Igor Gilitschenski

专题命中 可控生成 :diffusion(title);image editing(abstract);分类 cs.CV、cs.GR

Comments Project page: https://ihe-kaii.github.io/CTRL-D/

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.16689 2024-09-26 cs.CV cs.AI cs.GR cs.LG 81%

Layout-Corrector: Alleviating Layout Sticking Phenomenon in Discrete Diffusion Model

Shoma Iwai, Atsuki Osanai, Shunsuke Kitada, Shinichiro Omachi

专题命中 可控生成 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Accepted by ECCV2024, Project Page: https://iwa-shi.github.io/Layout-Corrector-Project-Page/

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.12960 2024-09-20 cs.CV cs.GR 81%

LVCD: Reference-based Lineart Video Colorization with Diffusion Models

Zhitong Huang, Mohan Zhang, Jing Liao

专题命中 可控生成 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Accepted by ACM Transactions on Graphics and SIGGRAPH Asia 2024. Project page: https://luckyhzt.github.io/lvcd

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.14101 2024-09-18 cs.CV cs.GR cs.LG 81%

Enhancing Image Layout Control with Loss-Guided Diffusion Models

Zakaria Patel, Kirill Serkh

专题命中 可控生成 :diffusion(title,abstract);分类 cs.CV、cs.GR

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.17414 2024-05-28 cs.CV cs.GR 81%

Collaborative Video Diffusion: Consistent Multi-video Generation with Camera Control

Zhengfei Kuang, Shengqu Cai, Hao He, Yinghao Xu, Hongsheng Li, Leonidas Guibas, Gordon Wetzstein

专题命中 可控生成 :diffusion(title,abstract);分类 cs.CV、cs.GR

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.15121 2024-04-24 cs.GR cs.AI cs.CV 81%

Taming Diffusion Probabilistic Models for Character Control

Rui Chen, Mingyi Shi, Shaoli Huang, Ping Tan, Taku Komura, Xuelin Chen

专题命中 可控生成 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Accepted by SIGGRAPH 2024 (Conference Track). Project page and source codes: https://aiganimation.github.io/CAMDM/

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.10615 2024-03-26 cs.CV cs.GR cs.LG 81%

LightIt: Illumination Modeling and Control for Diffusion Models

Peter Kocsis, Julien Philip, Kalyan Sunkavalli, Matthias Nießner, Yannick Hold-Geoffroy

专题命中 可控生成 :diffusion(title);image generation(abstract);分类 cs.CV、cs.GR

Comments Project page: https://peter-kocsis.github.io/LightIt/ Video: https://youtu.be/cCfSBD5aPLI

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.08778 2024-03-15 cs.CV cs.GR eess.IV 81%

Faster Projected GAN: Towards Faster Few-Shot Image Generation

Chuang Wang, Zhengping Li, Yuwen Hao, Lijun Wang, Xiaoxue Li

专题命中 可控生成 :image generation(title,abstract);分类 cs.CV、cs.GR

Comments 9 pages,7 figures,4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.02970 2023-12-06 cs.CV cs.AI cs.GR 81%

Alchemist: Parametric Control of Material Properties with Diffusion Models

Prafull Sharma, Varun Jampani, Yuanzhen Li, Xuhui Jia, Dmitry Lagun, Fredo Durand, William T. Freeman, Mark Matthews

专题命中 可控生成 :diffusion(title);text-to-image(abstract);分类 cs.CV、cs.GR

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.13153 2023-09-27 cs.CV cs.GR cs.LG 81%

Directed Diffusion: Direct Control of Object Placement through Attention Guidance

Wan-Duo Kurt Ma, J. P. Lewis, Avisek Lahiri, Thomas Leung, W. Bastiaan Kleijn

专题命中 可控生成 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Our project page: https://hohonu-vicml.github.io/DirectedDiffusion.Page

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.06725 2023-08-29 cs.CV cs.AI cs.MM 81%

CLE Diffusion: Controllable Light Enhancement Diffusion Model

Yuyang Yin, Dejia Xu, Chuangchuang Tan, Ping Liu, Yao Zhao, Yunchao Wei

专题命中 可控生成 :diffusion(title,abstract);分类 cs.CV、cs.MM

Comments Accepted In Proceedings of the 31st ACM International Conference on Multimedia (MM' 23)

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.04461 2023-05-10 cs.CV cs.GR 81%

Locally Attentional SDF Diffusion for Controllable 3D Shape Generation

Xin-Yang Zheng, Hao Pan, Peng-Shuai Wang, Xin Tong, Yang Liu, Heung-Yeung Shum

专题命中 可控生成 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Accepted to SIGGRAPH 2023 (Journal version)

Journal ref ACM Transactions on Graphics (SIGGRAPH), 42, 4 (August 2023), 13 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.04745 2023-05-09 cs.CV cs.GR 81%

Controllable Light Diffusion for Portraits

David Futschik, Kelvin Ritland, James Vecore, Sean Fanello, Sergio Orts-Escolano, Brian Curless, Daniel Sýkora, Rohit Pandey

专题命中 可控生成 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.08137 2023-03-15 cs.CV cs.GR 81%

LayoutDM: Discrete Diffusion Model for Controllable Layout Generation

Naoto Inoue, Kotaro Kikuchi, Edgar Simo-Serra, Mayu Otani, Kota Yamaguchi

专题命中 可控生成 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments To be published in CVPR2023, project page: https://cyberagentailab.github.io/layout-dm/

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.09778 2023-02-23 cs.CV cs.GR 81%

Composer: Creative and Controllable Image Synthesis with Composable Conditions

Lianghua Huang, Di Chen, Yu Liu, Yujun Shen, Deli Zhao, Jingren Zhou

专题命中 可控生成 :image synthesis(title);diffusion(abstract);分类 cs.CV、cs.GR

Comments Project page: https://damo-vilab.github.io/composer-page/

详情

展开后加载摘要…

URL PDF HTML 收藏
2105.14739 2023-01-11 cs.CV cs.GR 81%

Controllable Person Image Synthesis with Spatially-Adaptive Warped Normalization

Jichao Zhang, Aliaksandr Siarohin, Hao Tang, Enver Sangineto, Wei Wang, Humphrey Sh, Nicu Sebe

专题命中 可控生成 :image synthesis(title);image generation(abstract);分类 cs.CV、cs.GR

Comments 12 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.15796 2022-10-31 cs.CV cs.MM 81%

Layout Aware Inpainting for Automated Furniture Removal in Indoor Scenes

Prakhar Kulshreshtha, Konstantinos-Nektarios Lianos, Brian Pugh, Salma Jiddi

专题命中 可控生成 :inpainting(title,abstract);分类 cs.CV、cs.MM

Comments 6 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.11705 2022-05-25 cs.CV cs.AI cs.MM 81%

M6-Fashion: High-Fidelity Multi-modal Image Generation and Editing

Zhikang Li, Huiling Zhou, Shuai Bai, Peike Li, Chang Zhou, Hongxia Yang

专题命中 可控生成 :image generation(title,abstract);分类 cs.CV、cs.MM

Comments arXiv admin note: text overlap with arXiv:2105.14211

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.01068 2021-09-03 cs.CV cs.GR 81%

SLIDE: Single Image 3D Photography with Soft Layering and Depth-aware Inpainting

Varun Jampani, Huiwen Chang, Kyle Sargent, Abhishek Kar, Richard Tucker, Michael Krainin, Dominik Kaeser, William T. Freeman, David Salesin, Brian Curless, Ce Liu

专题命中 可控生成 :inpainting(title,abstract);分类 cs.CV、cs.GR

Comments ICCV 2021 (Oral); Project page: https://varunjampani.github.io/slide ; Video: https://www.youtube.com/watch?v=RQio7q-ueY8

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.13416 2021-06-30 cs.CV cs.GR 81%

Diversifying Semantic Image Synthesis and Editing via Class- and Layer-wise VAEs

Yuki Endo, Yoshihiro Kanamori

专题命中 可控生成 :image synthesis(title,abstract);分类 cs.CV、cs.GR

Comments Accepted to Pacific Graphics 2020, codes available at https://github.com/endo-yuki-t/DiversifyingSMIS

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.04644 2021-05-06 cs.CV cs.GR 81%

Efficient Semantic Image Synthesis via Class-Adaptive Normalization

Zhentao Tan, Dongdong Chen, Qi Chu, Menglei Chai, Jing Liao, Mingming He, Lu Yuan, Gang Hua, Nenghai Yu

专题命中 可控生成 :image synthesis(title,abstract);分类 cs.CV、cs.GR

Comments To appear at TPAMI 2021, code is available https://github.com/tzt101/CLADE.git

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.06878 2021-03-12 cs.CV cs.GR 81%

Diverse Semantic Image Synthesis via Probability Distribution Modeling

Zhentao Tan, Menglei Chai, Dongdong Chen, Jing Liao, Qi Chu, Bin Liu, Gang Hua, Nenghai Yu

专题命中 可控生成 :image synthesis(title,abstract);分类 cs.CV、cs.GR

Comments Accepted By CVPR 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.12861 2020-08-13 cs.CV cs.GR eess.IV 81%

SEAN: Image Synthesis with Semantic Region-Adaptive Normalization

Peihao Zhu, Rameen Abdal, Yipeng Qin, Peter Wonka

专题命中 可控生成 :image synthesis(title);image editing(abstract);分类 cs.CV、cs.GR

Comments Accepted as a CVPR 2020 oral paper. The interactive demo is available at https://youtu.be/0Vbj9xFgoUw

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.11758 2020-04-14 cs.CV cs.GR cs.LG eess.IV 81%

MixNMatch: Multifactor Disentanglement and Encoding for Conditional Image Generation

Yuheng Li, Krishna Kumar Singh, Utkarsh Ojha, Yong Jae Lee

专题命中 可控生成 :image generation(title,abstract);分类 cs.CV、cs.GR

Comments CVPR 2020 camera ready

Journal ref CVPR 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.12373 2019-08-30 cs.CV cs.GR cs.LG 81%

Diverse Image Synthesis from Semantic Layouts via Conditional IMLE

Ke Li, Tianhao Zhang, Jitendra Malik

专题命中 可控生成 :image synthesis(title,abstract);分类 cs.CV、cs.GR

Comments 18 pages, 16 figures; IEEE International Conference on Computer Vision (ICCV), 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.02526 2019-04-05 cs.GR cs.CV cs.LG stat.ML 81%

Constrained Generative Adversarial Networks for Interactive Image Generation

Eric Heim

专题命中 可控生成 :image generation(title,abstract);分类 cs.CV、cs.GR

Comments To Appear in the Proceedings of the 2019 Conference on Computer Vision and Pattern Recognition

详情

展开后加载摘要…

URL PDF HTML 收藏
1712.02225 2018-04-26 cs.CV cs.AI cs.MM stat.ML 81%

Pose-Normalized Image Generation for Person Re-identification

Xuelin Qian, Yanwei Fu, Tao Xiang, Wenxuan Wang, Jie Qiu, Yang Wu, Yu-Gang Jiang, Xiangyang Xue

专题命中 可控生成 :image generation(title,abstract);分类 cs.CV、cs.MM

Comments 10 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.06962 2025-06-17 cs.CV 80%

AR-RAG: Autoregressive Retrieval Augmentation for Image Generation

Jingyuan Qi, Zhiyang Xu, Qifan Wang, Lifu Huang

机构 * Virginia Tech(弗吉尼亚理工大学) Meta UC Davis(加州大学戴维斯分校)

专题命中 可控生成 :image generation(title,abstract);分类 cs.CV

Comments Image Generation, Retrieval Augmented Generation

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.13016 2025-03-21 cs.CV 80%

DiffPortrait3D: Controllable Diffusion for Zero-Shot Portrait View Synthesis

Yuming Gu, You Xie, Hongyi Xu, Guoxian Song, Yichun Shi, Di Chang, Jing Yang, Linjie Luo

专题命中 可控生成 :diffusion(title,abstract);分类 cs.CV

Journal ref https://openaccess.thecvf.com/content/CVPR2024/html/Gu_DiffPortrait3D_Controllable_Diffusion_for_Zero-Shot_Portrait_View_Synthesis_CVPR_2024_paper.html

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.03635 2025-02-14 cs.CV 80%

SSP-IR: Semantic and Structure Priors for Diffusion-based Realistic Image Restoration

Yuhong Zhang, Hengsheng Zhang, Zhengxue Cheng, Rong Xie, Li Song, Wenjun Zhang

专题命中 可控生成 :diffusion(title,abstract);分类 cs.CV

Comments To be published in IEEE TCSVT

Journal ref Y. Zhang, H. Zhang, Z. Cheng, R. Xie, L. Song and W. Zhang, "SSP-IR: Semantic and Structure Priors for Diffusion-based Realistic Image Restoration," in IEEE Transactions on Circuits and Systems for Video Technology, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏