arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 3474 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 3474 篇

2411.17472 2024-11-27 cs.CV cs.LG stat.ML 88%

Unlocking the Potential of Text-to-Image Diffusion with PAC-Bayesian Theory

Eric Hanchen Jiang, Yasi Zhang, Zhi Zhang, Yixin Wan, Andrew Lizarraga, Shufan Li, Ying Nian Wu

专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.16783 2024-11-27 cs.CV 88%

CoCoNO: Attention Contrast-and-Complete for Initial Noise Optimization in Text-to-Image Synthesis

Aravindan Sundaram, Ujjayan Pal, Abhimanyu Chauhan, Aishwarya Agarwal, Srikrishna Karanam

专题命中 文生图 :text-to-image(title,abstract);image synthesis(title);diffusion(abstract);分类 cs.CV

Comments 15 pages, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.01693 2024-11-26 cs.CV cs.AI 88%

HanDiffuser: Text-to-Image Generation With Realistic Hand Appearances

Supreeth Narasimhaswamy, Uttaran Bhattacharya, Xiang Chen, Ishita Dasgupta, Saayan Mitra, Minh Hoai

专题命中 文生图 :text-to-image(title,abstract);image generation(title);diffusion(abstract);分类 cs.CV

Comments Revisions: 1. Added a link to project page in the abstract, 2. Updated references and related work, 3. Fixed some grammatical errors

Journal ref In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2024, Seattle, Washington, USA

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.16394 2024-11-19 cs.LG cs.CV 88%

Skews in the Phenomenon Space Hinder Generalization in Text-to-Image Generation

Yingshan Chang, Yasi Zhang, Zhiyuan Fang, Yingnian Wu, Yonatan Bisk, Feng Gao

专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.05544 2024-11-11 cs.CV cs.LG 88%

Towards Lifelong Few-Shot Customization of Text-to-Image Diffusion

Nan Song, Xiaofeng Yang, Ze Yang, Guosheng Lin

专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.07605 2024-11-06 cs.CV cs.AI cs.LG 88%

Optimizing Negative Prompts for Enhanced Aesthetics and Fidelity in Text-To-Image Generation

Michael Ogezi, Ning Shi

专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.23962 2024-11-01 cs.CV cs.AI 88%

Image Synthesis with Class-Aware Semantic Diffusion Models for Surgical Scene Segmentation

Yihang Zhou, Rebecca Towning, Zaid Awad, Stamatia Giannarou

专题命中 文生图 :diffusion(title,abstract);image synthesis(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.12383 2024-10-29 cs.CV 88%

Reliable and Efficient Concept Erasure of Text-to-Image Diffusion Models

Chao Gong, Kai Chen, Zhipeng Wei, Jingjing Chen, Yu-Gang Jiang

专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV

Comments ECCV 2024 accepted

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.18615 2024-10-25 cs.CV cs.AI 88%

FairQueue: Rethinking Prompt Learning for Fair Text-to-Image Generation

Christopher T. H Teo, Milad Abdollahzadeh, Xinda Ma, Ngai-man Cheung

专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);分类 cs.CV

Comments Accepted in NeurIPS24

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.19599 2024-10-25 cs.CV cs.AI 88%

RealignDiff: Boosting Text-to-Image Diffusion Model with Coarse-to-fine Semantic Re-alignment

Zutao Jiang, Guian Fang, Jianhua Han, Guansong Lu, Hang Xu, Shengcai Liao, Xiaojun Chang, Xiaodan Liang

专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.17594 2024-10-24 cs.CV 88%

How to Continually Adapt Text-to-Image Diffusion Models for Flexible Customization?

Jiahua Dong, Wenqi Liang, Hongliu Li, Duzhen Zhang, Meng Cao, Henghui Ding, Salman Khan, Fahad Shahbaz Khan

专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV

Comments Accepted to NeurIPS2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.17255 2024-10-24 cs.CY cs.CV cs.LG 88%

Uncovering Regional Defaults from Photorealistic Forests in Text-to-Image Generation with DALL-E 2

Zilong Liu, Krzysztof Janowicz, Kitty Currier, Meilin Shi

专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);分类 cs.CV

Comments Accepted by the 16th Conference on Spatial Information Theory (COSIT 2024): https://cosit.ca

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.08551 2024-10-18 cs.CV cs.AI 88%

Context-Aware Full Body Anonymization using Text-to-Image Diffusion Models

Pascal Zwick, Kevin Roesch, Marvin Klemp, Oliver Bringmann

专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.16807 2024-10-18 cs.LG cs.CL cs.CV 88%

Beyond Thumbs Up/Down: Untangling Challenges of Fine-Grained Feedback for Text-to-Image Generation

Katherine M. Collins, Najoung Kim, Yonatan Bitton, Verena Rieser, Shayegan Omidshafiei, Yushi Hu, Sherol Chen, Senjuti Dutta, Minsuk Chang, Kimin Lee, Youwei Liang, Georgina Evans, Sahil Singla, Gang Li, Adrian Weller, Junfeng He, Deepak Ramachandran, Krishnamurthy Dj Dvijotham

专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.10437 2024-10-15 cs.CY cs.CV 88%

Towards Reliable Verification of Unauthorized Data Usage in Personalized Text-to-Image Diffusion Models

Boheng Li, Yanhao Wei, Yankai Fu, Zhenting Wang, Yiming Li, Jie Zhang, Run Wang, Tianwei Zhang

专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV

Comments To appear in the IEEE Symposium on Security & Privacy, May 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.13807 2024-10-14 cs.CV cs.AI cs.LG 88%

Editing Massive Concepts in Text-to-Image Diffusion Models

Tianwei Xiong, Yue Wu, Enze Xie, Yue Wu, Zhenguo Li, Xihui Liu

专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV

Comments Project page: https://silentview.github.io/EMCID/ . Code: https://github.com/SilentView/EMCID

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.07133 2024-10-11 cs.CV 88%

EvolveDirector: Approaching Advanced Text-to-Image Generation with Large Vision-Language Models

Rui Zhao, Hangjie Yuan, Yujie Wei, Shiwei Zhang, Yuchao Gu, Lingmin Ran, Xiang Wang, Zhangjie Wu, Junhao Zhang, Yingya Zhang, Mike Zheng Shou

专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.19967 2024-10-01 cs.CV 88%

Magnet: We Never Know How Text-to-Image Diffusion Models Work, Until We Learn How Vision-Language Models Function

Chenyi Zhuang, Ying Hu, Pan Gao

专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV

Comments Accepted to NeurIPS 2024. Code is available at https://github.com/I2-Multimedia-Lab/Magnet

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.16201 2024-09-26 cs.CV cs.AI cs.CL cs.LG 88%

Pre-trained Language Models Do Not Help Auto-regressive Text-to-Image Generation

Yuhui Zhang, Brandon McKinzie, Zhe Gan, Vaishaal Shankar, Alexander Toshev

专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);分类 cs.CV

Comments Published at EMNLP 2024 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.10476 2024-09-17 cs.CV 88%

SimInversion: A Simple Framework for Inversion-Based Text-to-Image Editing

Qi Qian, Haiyang Xu, Ming Yan, Juhua Hu

专题命中 文生图 :text-to-image(title);image editing(title);image generation(abstract);diffusion(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.08251 2024-09-13 cs.CV 88%

Dynamic Prompting of Frozen Text-to-Image Diffusion Models for Panoptic Narrative Grounding

Hongyu Li, Tianrui Hui, Zihan Ding, Jing Zhang, Bin Ma, Xiaoming Wei, Jizhong Han, Si Liu

专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV

Comments Accepted by ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.06493 2024-09-11 cs.CV cs.AI 88%

Elucidating Optimal Reward-Diversity Tradeoffs in Text-to-Image Diffusion Models

Rohit Jena, Ali Taghibakhshi, Sahil Jain, Gerald Shen, Nima Tajbakhsh, Arash Vahdat

专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.02429 2024-09-05 cs.CV 88%

Training-free Color-Style Disentanglement for Constrained Text-to-Image Synthesis

Aishwarya Agarwal, Srikrishna Karanam, Balaji Vasan Srinivasan

专题命中 文生图 :text-to-image(title,abstract);image synthesis(title);diffusion(abstract);分类 cs.CV

Comments 16 pages, 17 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.15721 2024-08-29 cs.CV 88%

Defending Text-to-image Diffusion Models: Surprising Efficacy of Textual Perturbations Against Backdoor Attacks

Oscar Chew, Po-Yi Lu, Jayden Lin, Hsuan-Tien Lin

专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV

Comments ECCV 2024 Workshop The Dark Side of Generative AIs and Beyond

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.08768 2024-08-23 cs.CV 88%

Local Conditional Controlling for Text-to-Image Diffusion Models

Yibo Zhao, Liang Peng, Yang Yang, Zekai Luo, Hengjia Li, Yao Chen, Zheng Yang, Xiaofei He, Wei Zhao, qinglin lu, Boxi Wu, Wei Liu

专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.08031 2024-08-20 cs.CV cs.AI cs.LG 88%

Latent Guard: a Safety Framework for Text-to-image Generation

Runtao Liu, Ashkan Khakzar, Jindong Gu, Qifeng Chen, Philip Torr, Fabio Pizzati

专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);分类 cs.CV

Comments This paper has been accepted to ECCV 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.02160 2024-08-06 cs.CV cs.CL 88%

ANNA: Abstractive Text-to-Image Synthesis with Filtered News Captions

Aashish Anantha Ramakrishnan, Sharon X. Huang, Dongwon Lee

专题命中 文生图 :text-to-image(title,abstract);image synthesis(title,abstract);分类 cs.CV

Comments To appear in the ACL 3rd Workshop on Advances in Language and Vision Research (ALVR), Bangkok, Thailand, August 2024, https://alvr-workshop.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.21032 2024-08-01 cs.CV cs.AI 88%

Safeguard Text-to-Image Diffusion Models with Human Feedback Inversion

Sanghyun Kim, Seohyeon Jung, Balhae Kim, Moonseok Choi, Jinwoo Shin, Juho Lee

专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV

Comments ECCV 2024. 56 pages, 24 figures. Caution: This paper contains discussions and examples related to harmful content, including text and images. Reader discretion is advised. Code is available at https://github.com/nannullna/safeguard-hfi

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.19996 2024-07-30 cs.CV cs.AI 88%

Reproducibility Study of "ITI-GEN: Inclusive Text-to-Image Generation"

Daniel Gallo Fernández, Răzvan-Andrei Matisan, Alejandro Monroy Muñoz, Janusz Partyka

专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);分类 cs.CV

Comments Accepted to TMLR, see https://openreview.net/forum?id=d3Vj360Wi2

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.01807 2024-07-30 cs.CV 88%

ViewDiff: 3D-Consistent Image Generation with Text-to-Image Models

Lukas Höllein, Aljaž Božič, Norman Müller, David Novotny, Hung-Yu Tseng, Christian Richardt, Michael Zollhöfer, Matthias Nießner

专题命中 文生图 :text-to-image(title,abstract);image generation(title);diffusion(abstract);分类 cs.CV

Comments Accepted to CVPR 2024, project page: https://lukashoel.github.io/ViewDiff/, video: https://www.youtube.com/watch?v=SdjoCqHzMMk, code: https://github.com/facebookresearch/ViewDiff

详情

展开后加载摘要…

URL PDF HTML 收藏