arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 86504 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 3480 篇

2305.12716 2024-02-16 cs.CV 81%

The CLIP Model is Secretly an Image-to-Prompt Converter

Yuxuan Ding, Chunna Tian, Haoxuan Ding, Lingqiao Liu

专题命中 文生图 :image generation(abstract);text-to-image(abstract);diffusion(abstract);image editing(abstract)

Comments Accepted by NeurIPS 2023, 21 pages, 28 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.00813 2023-06-02 cs.CV 81%

UniDiff: Advancing Vision-Language Models with Generative and Discriminative Learning

Xiao Dong, Runhui Huang, Xiaoyong Wei, Zequn Jie, Jianxing Yu, Jian Yin, Xiaodan Liang

专题命中 文生图 :image generation(abstract);text-to-image(abstract);diffusion(abstract);image synthesis(abstract)

Comments NA

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.03502 2021-12-08 cs.LG cs.CV 81%

A Generic Approach for Enhancing GANs by Regularized Latent Optimization

Yufan Zhou, Chunyuan Li, Changyou Chen, Jinhui Xu

专题命中 文生图 :image generation(abstract);text-to-image(abstract);image editing(abstract);inpainting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.15740 2026-08-17 cs.CV cs.AI cs.LG cs.MM 版本更新 81%

Debiasing Text-to-Image Evaluation via Implicit Cultural Alignment Reward Modeling

通过隐式文化对齐奖励建模消除文本到图像评估中的偏差

Bo-An Chang, Yu-Chih Chen

机构 * National Tsing Hua University(国立清华大学) National Yang Ming Chiao Tung University(国立阳明交通大学)

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV、cs.MM

AI总结 研究文本到图像评估中文化真实性问题,提出基于轻量级多模态大语言模型构建的隐式文化对齐奖励模型,集成隐式文化探测器与跳跃连接交叉注意力机制,实验证明该模型准确率高、速度快,能为偏好优化管道提供有效信号。

Comments 16 pages, 2 figures, ECCV 2026 Workshop FAILED

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.06074 2026-04-08 cs.CV cs.AI cs.MM 81%

Graph-PiT: Enhancing Structural Coherence in Part-Based Image Synthesis via Graph Priors

Graph-PiT:通过图先验增强基于部分的图像合成的结构一致性

Junbin Zhang, Meng Cao, Feng Tan, Yikai Lin, Yuexian Zou

机构 * Guangdong Provincial Key Laboratory of Ultra High Definition Immersive Media Technology(广东省超高清沉浸式媒体技术重点实验室) Shenzhen Graduate School, Peking University, China(北京大学深圳研究生院) Department of Computer Vision, Mohamed bin Zayed University of Artificial Intelligence, Abu Dhabi, United Arab Emirates(穆罕默德·本·扎耶德人工智能大学计算机视觉系)

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV、cs.MM

AI总结 Graph-PiT通过图先验建模视觉组件的结构依赖,提升部分基于图像合成的结构一致性,同时保持与原始IP-Prior流程的兼容性。

Comments 11 pages, 5 figures, Accepted by ICME 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.20201 2026-03-24 cs.MM cs.CV cs.CY 81%

FIGURA: A Modular Prompt Engineering Method for Artistic Figure Photography in Safety-Filtered Text-to-Image Models

FIGURA:一种用于安全过滤文本到图像模型中艺术人物摄影的模块化提示工程方法

Luca Cazzaniga

机构 * Independent Researcher(独立研究员)

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV、cs.MM

AI总结 本文提出FIGURA方法,通过模块化提示工程系统解决安全过滤文本到图像模型中艺术人物摄影的限制问题,通过200+测试验证其有效性,揭示安全过滤机制的运作原理并提供系统解决方案。

Comments 10 pages, 6 tables. Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.01720 2025-10-14 cs.CV cs.GR cs.LG 81%

Generating Multi-Image Synthetic Data for Text-to-Image Customization

Nupur Kumari, Xi Yin, Jun-Yan Zhu, Ishan Misra, Samaneh Azadi

机构 * Carnegie Mellon University(卡内基梅隆大学) Meta

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV、cs.GR

Comments ICCV 2025. Project webpage: https://www.cs.cmu.edu/~syncd-project/

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.17062 2025-04-25 cs.GR cs.CV 81%

ePBR: Extended PBR Materials in Image Synthesis

Yu Guo, Zhiqiang Lao, Xiyun Song, Yubin Zhou, Zongfang Lin, Heather Yu

机构 * Futurewei Technologies(未来科技公司)

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV、cs.GR

Comments 8 pages without references, 7 figures, accepted in CVPRW 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.05606 2025-03-11 cs.CV cs.MM 81%

CustomContrast: A Multilevel Contrastive Perspective For Subject-Driven Text-to-Image Customization

Nan Chen, Mengqi Huang, Zhuowei Chen, Yang Zheng, Lei Zhang, Zhendong Mao

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV、cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.01536 2024-10-29 cs.CV cs.GR cs.LG 81%

Customizing Text-to-Image Models with a Single Image Pair

Maxwell Jones, Sheng-Yu Wang, Nupur Kumari, David Bau, Jun-Yan Zhu

专题命中 文生图 :text-to-image(title);diffusion(abstract);分类 cs.CV、cs.GR

Comments project page: https://paircustomization.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.10695 2024-10-23 cs.CV cs.AI cs.GR 81%

Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models

Bingchen Liu, Ehsan Akhgari, Alexander Visheratin, Aleks Kamko, Linmiao Xu, Shivam Shrirao, Chase Lambert, Joao Souza, Suhail Doshi, Daiqing Li

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV、cs.GR

Comments Project page: https://playground.com/pg-v3

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.14844 2024-10-22 cs.CV cs.CE cs.GR 81%

SYNOSIS: Image synthesis pipeline for machine vision in metal surface inspection

Juraj Fulir, Natascha Jeziorski, Lovro Bosnar, Hans Hagen, Claudia Redenbach, Petra Gospodnetić, Tobias Herrfurth, Marcus Trost, Thomas Gischkat

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV、cs.GR

Comments Initial preprint, 21 pages, 21 figures, 6 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.17274 2024-07-25 cs.MM cs.AI cs.CV 81%

Revolutionizing Text-to-Image Retrieval as Autoregressive Token-to-Voken Generation

Yongqi Li, Hongru Cai, Wenjie Wang, Leigang Qu, Yinwei Wei, Wenjie Li, Liqiang Nie, Tat-Seng Chua

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV、cs.MM

Comments Work in progress

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.01644 2024-06-06 cs.CV cs.AI cs.GR 81%

Key-Locked Rank One Editing for Text-to-Image Personalization

Yoad Tewel, Rinon Gal, Gal Chechik, Yuval Atzmon

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV、cs.GR

Comments Accepted to SIGGRAPH 2023. Project page is in https://research.nvidia.com/labs/par/Perfusion/

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.00057 2024-04-03 cs.CR cs.AI cs.CV cs.MM 81%

VA3: Virtually Assured Amplification Attack on Probabilistic Copyright Protection for Text-to-Image Generative Models

Xiang Li, Qianli Shen, Kenji Kawaguchi

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV、cs.MM

Comments 18 pages, 9 figures. Accept to CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.09911 2024-03-29 cs.CV cs.MM 81%

Noisy-Correspondence Learning for Text-to-Image Person Re-identification

Yang Qin, Yingke Chen, Dezhong Peng, Xi Peng, Joey Tianyi Zhou, Peng Hu

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV、cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.06105 2024-01-12 cs.CV cs.CL cs.GR cs.LG 81%

PALP: Prompt Aligned Personalization of Text-to-Image Models

Moab Arar, Andrey Voynov, Amir Hertz, Omri Avrahami, Shlomi Fruchter, Yael Pritch, Daniel Cohen-Or, Ariel Shamir

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV、cs.GR

Comments Project page available at https://prompt-aligned.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.11210 2023-10-18 cs.CV cs.MM 81%

Learning Comprehensive Representations with Richer Self for Text-to-Image Person Re-Identification

Shuanglin Yan, Neng Dong, Jun Liu, Liyan Zhang, Jinhui Tang

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV、cs.MM

Comments Accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.14267 2023-10-04 cs.CV cs.AI cs.GR 81%

A Survey on Deep Generative 3D-aware Image Synthesis

Weihao Xia, Jing-Hao Xue

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV、cs.GR

Comments Accepted to ACM Computing Surveys. Project page: https://weihaox.github.io/3D-aware-Gen

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.10388 2023-09-20 cs.CV cs.GR 81%

SideGAN: 3D-Aware Generative Model for Improved Side-View Image Synthesis

Kyungmin Jo, Wonjoon Jin, Jaegul Choo, Hyunjoon Lee, Sunghyun Cho

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV、cs.GR

Comments International Conference on Computer Vision (ICCV) 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.08509 2023-05-02 cs.CV cs.GR cs.LG 81%

3D-aware Conditional Image Synthesis

Kangle Deng, Gengshan Yang, Deva Ramanan, Jun-Yan Zhu

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV、cs.GR

Comments Project Page: https://www.cs.cmu.edu/~pix2pix3D/

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.12503 2022-11-24 cs.CL cs.CV cs.LG cs.MM 81%

Is the Elephant Flying? Resolving Ambiguities in Text-to-Image Generative Models

Ninareh Mehrabi, Palash Goyal, Apurv Verma, Jwala Dhamala, Varun Kumar, Qian Hu, Kai-Wei Chang, Richard Zemel, Aram Galstyan, Rahul Gupta

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV、cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.11832 2022-03-23 cs.CV cs.MM 81%

Cross-View Panorama Image Synthesis

Songsong Wu, Hao Tang, Xiao-Yuan Jing, Haifeng Zhao, Jianjun Qian, Nicu Sebe, Yan Yan

专题命中 文生图 :image synthesis(title);image generation(abstract);分类 cs.CV、cs.MM

Comments Accepted to IEEE Transactions on Multimedia

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.00659 2022-02-02 cs.CV cs.GR 81%

Stay Positive: Non-Negative Image Synthesis for Augmented Reality

Katie Luo, Guandao Yang, Wenqi Xian, Harald Haraldsson, Bharath Hariharan, Serge Belongie

专题命中 文生图 :image synthesis(title);image generation(abstract);分类 cs.CV、cs.GR

Journal ref Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2021, pp. 10050-10060

详情

展开后加载摘要…

URL PDF HTML 收藏
2111.03384 2021-11-08 cs.CV cs.GR 81%

Seamless Satellite-image Synthesis

Jialin Zhu, Tom Kelly

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV、cs.GR

Comments 12 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.05783 2021-09-14 cs.CV cs.GR 81%

The State of the Art when using GPUs in Devising Image Generation Methods Using Deep Learning

Yasuko Kawahata

专题命中 文生图 :image generation(title);image synthesis(abstract);分类 cs.CV、cs.GR

Comments 13 pages

Journal ref Research Report (2016/03)

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.00926 2021-04-07 cs.CV cs.GR 81%

pi-GAN: Periodic Implicit Generative Adversarial Networks for 3D-Aware Image Synthesis

Eric R. Chan, Marco Monteiro, Petr Kellnhofer, Jiajun Wu, Gordon Wetzstein

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV、cs.GR

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.13024 2020-09-01 cs.CV cs.AI cs.LG cs.MM 81%

Dual Attention GANs for Semantic Image Synthesis

Hao Tang, Song Bai, Nicu Sebe

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV、cs.MM

Comments Accepted to ACM MM 2020, camera ready (9 pages) + supplementary (10 pages)

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.03590 2020-04-09 cs.CV cs.GR cs.LG cs.NE eess.IV 81%

Multimodal Image Synthesis with Conditional Implicit Maximum Likelihood Estimation

Ke Li, Shichong Peng, Tianhao Zhang, Jitendra Malik

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV、cs.GR

Comments To appear in International Journal of Computer Vision (IJCV). arXiv admin note: text overlap with arXiv:1811.12373

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.12356 2019-04-30 cs.CV cs.GR 81%

Deferred Neural Rendering: Image Synthesis using Neural Textures

Justus Thies, Michael Zollhöfer, Matthias Nießner

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV、cs.GR

Comments Video: https://youtu.be/z-pVip6WeyY SIGGRAPH 2019

详情

展开后加载摘要…

URL PDF HTML 收藏