arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 3480 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 3480 篇

2403.15330 2024-03-25 cs.CV 79%

Selectively Informative Description can Reduce Undesired Embedding Entanglements in Text-to-Image Personalization

Jimyeong Kim, Jungwon Park, Wonjong Rhee

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

Comments Published at CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.11092 2024-03-19 cs.CL cs.AI cs.CV cs.CY eess.IV 79%

Lost in Translation? Translation Errors and Challenges for Fair Assessment of Text-to-Image Models on Multilingual Concepts

Michael Saxon, Yiran Luo, Sharon Levy, Chitta Baral, Yezhou Yang, William Yang Wang

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

Comments NAACL 2024 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.06835 2024-03-12 cs.CV cs.AI cs.CL 79%

Medical Image Synthesis via Fine-Grained Image-Text Alignment and Anatomy-Pathology Prompting

Wenting Chen, Pengyu Wang, Hui Ren, Lichao Sun, Quanzheng Li, Yixuan Yuan, Xiang Li

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.00483 2024-03-04 cs.CV 79%

RealCustom: Narrowing Real Text Word for Real-Time Open-Domain Text-to-Image Customization

Mengqi Huang, Zhendong Mao, Mingcong Liu, Qian He, Yongdong Zhang

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

Comments Accepted by CVPR2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.03661 2024-02-22 cs.CV 79%

Robustness-Guided Image Synthesis for Data-Free Quantization

Jianhong Bai, Yuchen Yang, Huanpeng Chu, Hualiang Wang, Zuozhu Liu, Ruizhe Chen, Xiaoxuan He, Lianrui Mu, Chengfei Cai, Haoji Hu

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments Accepted at AAAI 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.09036 2024-02-15 cs.CV 79%

Can Text-to-image Model Assist Multi-modal Learning for Visual Recognition with Visual Modality Missing?

Tiantian Feng, Daniel Yang, Digbalay Bose, Shrikanth Narayanan

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.09944 2024-02-06 cs.LG cs.AI cs.CV cs.CY 79%

DiffusionWorldViewer: Exposing and Broadening the Worldview Reflected by Generative Text-to-Image Models

Zoe De Simone, Angie Boggust, Arvind Satyanarayan, Ashia Wilson

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

Comments 20 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.07294 2024-01-24 cs.MM 79%

Probing Commonsense Reasoning Capability of Text-to-Image Generative Models via Non-visual Description

Mianzhi Pan, Jianfei Li, Mingyue Yu, Zheng Ma, Kanzhi Cheng, Jianbing Zhang, Jiajun Chen

专题命中 文生图 :text-to-image(title,abstract);分类 cs.MM

Comments It is an incomplete work

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.04241 2024-01-10 cs.CV 79%

Data-Agnostic Face Image Synthesis Detection Using Bayesian CNNs

Roberto Leyva, Victor Sanchez, Gregory Epiphaniou, Carsten Maple

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.10373 2024-01-08 eess.IV cs.CV q-bio.QM q-bio.TO 79%

A SSIM Guided cGAN Architecture For Clinically Driven Generative Image Synthesis of Multiplexed Spatial Proteomics Channels

Jillur Rahman Saurav, Mohammad Sadegh Nasr, Paul Koomey, Michael Robben, Manfred Huber, Jon Weidanz, Bríd Ryan, Eytan Ruppin, Peng Jiang, Jacob M. Luber

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Journal ref 2023 IEEE CIBCB, Eindhoven, Netherlands, 2023, pp. 1-8

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.02173 2024-01-05 cs.CV cs.AI 79%

Prompt Decoupling for Text-to-Image Person Re-identification

Weihao Li, Lei Tan, Pingyang Dai, Yan Zhang

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.00314 2024-01-02 eess.IV cs.CV cs.LG cs.NE 79%

GAN-GA: A Generative Model based on Genetic Algorithm for Medical Image Generation

M. AbdulRazek, G. Khoriba, M. Belal

专题命中 文生图 :image generation(title);image synthesis(abstract);分类 cs.CV

Comments 10 pages, 2 figures. Abstract published in Frontiers in Medical Technology, presented at the 27th Conference on Medical Image Understanding and Analysis 2023. DOI: 10.3389/978-2-8325-1231-9. URL: https://doi.org/10.3389/978-2-8325-1231-9

Journal ref 27th Conference on Medical Image Understanding and Analysis 2023, Frontiers, 2023, pp. 30-39

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.13053 2023-12-21 cs.CV cs.CR 79%

Quantifying Bias in Text-to-Image Generative Models

Jordan Vice, Naveed Akhtar, Richard Hartley, Ajmal Mian

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

Comments main manuscript = 9 pages, 6 tables, 4 figures. Supplementary material = 15 pages, 13 tables, 14 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.04779 2023-12-11 eess.IV cs.CV cs.LG 79%

Image Synthesis-based Late Stage Cancer Augmentation and Semi-Supervised Segmentation for MRI Rectal Cancer Staging

Saeko Sasuga, Akira Kudo, Yoshiro Kitamura, Satoshi Iizuka, Edgar Simo-Serra, Atsushi Hamabe, Masayuki Ishii, Ichiro Takemasa

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments 10 pages, 7 figures, Accepted to Data Augmentation, Labeling, and Imperfections (DALI) at MICCAI 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.01745 2023-12-05 cs.CV 79%

Cross-Modal Adaptive Dual Association for Text-to-Image Person Retrieval

Dixuan Lin, Yixing Peng, Jingke Meng, Wei-Shi Zheng

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.00570 2023-12-04 cs.CV 79%

Generative models for visualising abstract social processes: Guiding streetview image synthesis of StyleGAN2 with indices of deprivation

Aleksi Knuutila

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments 10 pages, 3 figures, 1 table, associated website with interactive interface at http://site.knuutila.net/thisinequalitydoesnotexist

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.09702 2023-11-15 eess.IV cs.CV cs.LG 79%

Illumination Variation Correction Using Image Synthesis For Unsupervised Domain Adaptive Person Re-Identification

Jiaqi Guo, Amy R. Reibman, Edward J. Delp

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments 10 pages, 5 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.06329 2023-11-14 cs.CV cs.AI cs.CL cs.LG eess.IV 79%

A Survey of AI Text-to-Image and AI Text-to-Video Generators

Aditi Singh

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

Comments 4 pages, 2 tables, 4th International Conference on Artificial Intelligence, Robotics and Control (AIRC 2023)

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.01755 2023-10-26 cs.CV cs.AI cs.CL 79%

Training Priors Predict Text-To-Image Model Performance

Charles Lovering, Ellie Pavlick

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.01508 2023-10-10 cs.LG cs.CR cs.CV 79%

Circumventing Concept Erasure Methods For Text-to-Image Generative Models

Minh Pham, Kelly O. Marshall, Niv Cohen, Govind Mittal, Chinmay Hegde

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.12491 2023-10-02 cs.CV cs.LG eess.IV 79%

Deformation equivariant cross-modality image synthesis with paired non-aligned training data

Joel Honkamaa, Umair Khan, Sonja Koivukoski, Mira Valkonen, Leena Latonen, Pekka Ruusuvuori, Pekka Marttinen

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Journal ref Medical Image Analysis 90 (2023): 102940

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.13216 2023-09-26 cs.CV cs.AI cs.HC cs.RO 79%

MISFIT-V: Misaligned Image Synthesis and Fusion using Information from Thermal and Visual

Aadhar Chauhan, Isaac Remy, Danny Broyles, Karen Leung

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.06997 2023-09-26 eess.IV cs.CV 79%

Cross-Modality Neuroimage Synthesis: A Survey

Guoyang Xie, Yawen Huang, Jinbao Wang, Jiayi Lyu, Feng Zheng, Yefeng Zheng, Yaochu Jin

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.04511 2023-09-20 eess.IV cs.AI cs.CV 79%

Systematic Review of Techniques in Brain Image Synthesis using Deep Learning

Shubham Singh, Ammar Ranapurwala, Mrunal Bewoor, Sheetal Patil, Satyam Rai

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments 8 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.01420 2023-09-06 cs.CV 79%

Unified Pre-training with Pseudo Texts for Text-To-Image Person Re-identification

Zhiyin Shao, Xinyu Zhang, Changxing Ding, Jian Wang, Jingdong Wang

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

Comments accepted by ICCV 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.13240 2023-09-04 cs.CV 79%

Contrastive Image Synthesis and Self-supervised Feature Adaptation for Cross-Modality Biomedical Image Segmentation

Xinrong Hu, Corey Wang, Yiyu Shi

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.14066 2023-08-30 eess.IV cs.CV 79%

Bi-Modality Medical Image Synthesis Using Semi-Supervised Sequential Generative Adversarial Networks

Xin Yang, Yi Lin, Zhiwei Wang, Xin Li, Kwang-Ting Cheng

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.13592 2023-08-25 cs.CV 79%

Multimodal Image Synthesis and Editing: The Generative AI Era

Fangneng Zhan, Yingchen Yu, Rongliang Wu, Jiahui Zhang, Shijian Lu, Lingjie Liu, Adam Kortylewski, Christian Theobalt, Eric Xing

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments TPAMI 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.05921 2023-08-14 cs.CV 79%

BATINet: Background-Aware Text to Image Synthesis and Manipulation Network

Ryugo Morita, Zhiqiang Zhang, Jinjia Zhou

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments Accepted to ICIP2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.05242 2023-08-11 cs.CV cs.AI 79%

Vector quantization loss analysis in VQGANs: a single-GPU ablation study for image-to-image synthesis

Luv Verma, Varun Mohan

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments 16 pages, 18 figures

详情

展开后加载摘要…

URL PDF HTML 收藏