arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 86504 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 3480 篇

2409.06385 2025-02-11 cs.CV 79%

AMNS: Attention-Weighted Selective Mask and Noise Label Suppression for Text-to-Image Person Retrieval

Runqing Zhang, Xue Zhou

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.10848 2025-02-11 cs.CV 79%

Perception-guided Jailbreak against Text-to-Image Models

Yihao Huang, Le Liang, Tianlin Li, Xiaojun Jia, Run Wang, Weikai Miao, Geguang Pu, Yang Liu

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

Comments 9 pages, accepted by AAAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.14774 2025-02-07 cs.LG cs.CL cs.CV 79%

Evaluating Numerical Reasoning in Text-to-Image Models

Ivana Kajić, Olivia Wiles, Isabela Albuquerque, Matthias Bauer, Su Wang, Jordi Pont-Tuset, Aida Nematzadeh

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.03420 2025-02-06 cs.CV 79%

Can Text-to-Image Generative Models Accurately Depict Age? A Comparative Study on Synthetic Portrait Generation and Age Estimation

Alexey A. Novikov, Miroslav Vranka, François David, Artem Voronin

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.11962 2025-01-31 cs.AI cs.CR cs.CV cs.LG 79%

©Plug-in Authorization for Human Content Copyright Protection in Text-to-Image Model

Chao Zhou, Huishuai Zhang, Jiang Bian, Weiming Zhang, Nenghai Yu

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

Comments 23 pages, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.19510 2025-01-28 cs.CV cs.LG 79%

Retrieval-guided Cross-view Image Synthesis

Hongji Yang, Yiru Li, Yingying Zhu

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.11815 2025-01-23 cs.CV 79%

CogMorph: Cognitive Morphing Attacks for Text-to-Image Models

Zonglei Jing, Zonghao Ying, Le Wang, Siyuan Liang, Aishan Liu, Xianglong Liu, Dacheng Tao

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.06863 2025-01-22 cs.CV 79%

Beyond Aesthetics: Cultural Competence in Text-to-Image Models

Nithish Kannen, Arif Ahmad, Marco Andreetto, Vinodkumar Prabhakaran, Utsav Prabhu, Adji Bousso Dieng, Pushpak Bhattacharyya, Shachi Dave

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

Comments NeurIPS 2024 camera-ready version

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.06770 2025-01-20 cs.CV 79%

SuperNeRF-GAN: A Universal 3D-Consistent Super-Resolution Framework for Efficient and Enhanced 3D-Aware Image Synthesis

Peng Zheng, Linzhi Huang, Yizhou Yu, Yi Chang, Yilin Wang, Rui Ma

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.09059 2025-01-20 cs.CL cs.AI cs.CV 79%

Text-guided Image Restoration and Semantic Enhancement for Text-to-Image Person Retrieval

Delong Liu, Haiwen Li, Zhicheng Zhao, Yuan Dong

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

Comments The paper was withdrawn due to a dispute among the authors regarding the content of the article

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.21035 2025-01-17 cs.CV 79%

Direct Unlearning Optimization for Robust and Safe Text-to-Image Models

Yong-Hyun Park, Sangdoo Yun, Jin-Hwa Kim, Junho Kim, Geonhui Jang, Yonghyun Jeong, Junghyo Jo, Gayoung Lee

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

Comments This paper has been accepted for NeurIPS 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.06356 2025-01-14 eess.IV cs.AI cs.CV 79%

Ultrasound Image Synthesis Using Generative AI for Lung Ultrasound Detection

Yu-Cheng Chou, Gary Y. Li, Li Chen, Mohsen Zahiri, Naveen Balaraju, Shubham Patil, Bryson Hicks, Nikolai Schnittke, David O. Kessler, Jeffrey Shupp, Maria Parker, Cristiana Baloescu, Christopher Moore, Cynthia Gregory, Kenton Gregory, Balasundar Raju, Jochen Kruecker, Alvin Chen

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments Accepted by ISBI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.12104 2025-01-03 cs.CV cs.CL cs.LG 79%

Ethical-Lens: Curbing Malicious Usages of Open-Source Text-to-Image Models

Yuzhu Cai, Sheng Yin, Yuxi Wei, Chenxin Xu, Weibo Mao, Felix Juefei-Xu, Siheng Chen, Yanfeng Wang

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

Comments 51 pages, 15 figures, 32 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.19531 2024-12-30 cs.CV cs.AI 79%

Is Your Text-to-Image Model Robust to Caption Noise?

Weichen Yu, Ziyan Yang, Shanchuan Lin, Qi Zhao, Jianyi Wang, Liangke Gui, Matt Fredrikson, Lu Jiang

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.16879 2024-12-19 cs.CV cs.AI 79%

Image Synthesis under Limited Data: A Survey and Taxonomy

Mengping Yang, Zhe Wang

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments 230 references, 25 pages. GitHub: https://github.com/kobeshegu/awesome-few-shot-generation

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.12774 2024-12-18 cs.CV cs.CY 79%

A Framework for Critical Evaluation of Text-to-Image Models: Integrating Art Historical Analysis, Artistic Exploration, and Critical Prompt Engineering

Amalia Foka

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.07196 2024-12-17 cs.CV 79%

Fine-grained Text to Image Synthesis

Xu Ouyang, Ying Chen, Kaiyue Zhu, Gady Agam

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.04324 2024-12-06 physics.med-ph cs.CV 79%

Multi-Subject Image Synthesis as a Generative Prior for Single-Subject PET Image Reconstruction

George Webber, Yuya Mizuno, Oliver D. Howes, Alexander Hammers, Andrew P. King, Andrew J. Reader

专题命中 文生图 :image synthesis(title);diffusion(abstract);分类 cs.CV

Comments 2 pages, 3 figures. Accepted as a poster presentation at IEEE NSS MIC RTSD 2024 (submitted May 2024; accepted July 2024; presented Nov 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.03957 2024-12-06 cs.CV cs.AI 79%

A Framework For Image Synthesis Using Supervised Contrastive Learning

Yibin Liu, Jianyu Zhang, Li Zhang, Shijian Li, Gang Pan

专题命中 文生图 :image synthesis(title);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.09011 2024-11-26 eess.IV cs.CV 79%

The Brain Tumor Segmentation (BraTS) Challenge 2023: Brain MR Image Synthesis for Tumor Segmentation (BraSyn)

Hongwei Bran Li, Gian Marco Conte, Qingqiao Hu, Syed Muhammad Anwar, Florian Kofler, Ivan Ezhov, Koen van Leemput, Marie Piraud, Maria Diaz, Byrone Cole, Evan Calabrese, Jeff Rudie, Felix Meissen, Maruf Adewole, Anastasia Janas, Anahita Fathi Kazerooni, Dominic LaBella, Ahmed W. Moawad, Keyvan Farahani, James Eddy, Timothy Bergquist, Verena Chung, Russell Takeshi Shinohara, Farouk Dako, Walter Wiggins, Zachary Reitman, Chunhao Wang, Xinyang Liu, Zhifan Jiang, Ariana Familiar, Elaine Johanson, Zeke Meier, Christos Davatzikos, John Freymann, Justin Kirby, Michel Bilello, Hassan M. Fathallah-Shaykh, Roland Wiest, Jan Kirschke, Rivka R. Colen, Aikaterini Kotrotsou, Pamela Lamontagne, Daniel Marcus, Mikhail Milchenko, Arash Nazeri, Marc André Weber, Abhishek Mahajan, Suyash Mohan, John Mongan, Christopher Hess, Soonmee Cha, Javier Villanueva, Meyer Errol Colak, Priscila Crivellaro, Andras Jakab, Jake Albrecht, Udunna Anazodo, Mariam Aboian, Thomas Yu, Verena Chung, Timothy Bergquist, James Eddy, Jake Albrecht, Ujjwal Baid, Spyridon Bakas, Marius George Linguraru, Bjoern Menze, Juan Eugenio Iglesias, Benedikt Wiestler

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments Technical report of BraSyn

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.06959 2024-11-12 cs.CV cs.AI 79%

ENAT: Rethinking Spatial-temporal Interactions in Token-based Image Synthesis

Zanlin Ni, Yulin Wang, Renping Zhou, Yizeng Han, Jiayi Guo, Zhiyuan Liu, Yuan Yao, Gao Huang

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments Accepted by NeurIPS2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.09952 2024-11-05 cs.CV cs.CL cs.LG 79%

BiVLC: Extending Vision-Language Compositionality Evaluation with Text-to-Image Retrieval

Imanol Miranda, Ander Salaberria, Eneko Agirre, Gorka Azkune

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

Comments Accepted to NeurIPS 24 Datasets and Benchmarks Track; Project page at: https://imirandam.github.io/BiVLC_project_page/

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.04251 2024-11-01 cs.CV cs.AI cs.CL 79%

Who Evaluates the Evaluations? Objectively Scoring Text-to-Image Prompt Coherence Metrics with T2IScoreScore (TS2)

Michael Saxon, Fatima Jahara, Mahsa Khoshnoodi, Yujie Lu, Aditya Sharma, William Yang Wang

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

Comments NeurIPS 2024 Spotlight

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.01446 2024-10-31 cs.CV 79%

GuardT2I: Defending Text-to-Image Models from Adversarial Prompts

Yijun Yang, Ruiyuan Gao, Xiao Yang, Jianyuan Zhong, Qiang Xu

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

Comments NeurIPS2024 Poster

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.17691 2024-10-24 eess.IV cs.CV q-bio.NC 79%

Longitudinal Causal Image Synthesis

Yujia Li, Han Li, ans S. Kevin Zhou

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.14122 2024-10-18 cs.CV cs.CR 79%

SurrogatePrompt: Bypassing the Safety Filter of Text-to-Image Models via Substitution

Zhongjie Ba, Jieming Zhong, Jiachen Lei, Peng Cheng, Qinglong Wang, Zhan Qin, Zhibo Wang, Kui Ren

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

Comments To appear in the the 31st ACM Conference on Computer and Communications Security (CCS)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.08608 2024-10-14 cs.CV cs.AI cs.LG 79%

Text-To-Image with Generative Adversarial Networks

Mehrshad Momen-Tayefeh

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.18815 2024-10-01 cs.CV 79%

IMMA: Immunizing text-to-image Models against Malicious Adaptation

Amber Yijia Zheng, Raymond A. Yeh

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

Comments ECCV 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.16818 2024-09-26 eess.IV cs.CV 79%

Towards General Text-guided Image Synthesis for Customized Multimodal Brain MRI Generation

Yulin Wang, Honglin Xiong, Kaicong Sun, Shuwei Bai, Ling Dai, Zhongxiang Ding, Jiameng Liu, Qian Wang, Qian Liu, Dinggang Shen

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments 23 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.12010 2024-09-19 cs.CV 79%

ChefFusion: Multimodal Foundation Model Integrating Recipe and Food Image Generation

Peiyu Li, Xiaobao Huang, Yijun Tian, Nitesh V. Chawla

专题命中 文生图 :image generation(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏