arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 86504 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 3480 篇

2409.09969 2024-09-17 cs.CV eess.IV 79%

2S-ODIS: Two-Stage Omni-Directional Image Synthesis by Geometric Distortion Correction

Atsuya Nakata, Takao Yamanaka

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments ECCV2024 https://github.com/islab-sophia/2S-ODIS

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.09593 2024-09-17 cs.CV 79%

One-Shot Learning for Pose-Guided Person Image Synthesis in the Wild

Dongqi Fan, Tao Chen, Mingjie Wang, Rui Ma, Qiang Tang, Zili Yi, Qian Wang, Liang Chang

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.09427 2024-09-17 cs.MM 79%

Prototypical Prompting for Text-to-image Person Re-identification

Shuanglin Yan, Jun Liu, Neng Dong, Liyan Zhang, Jinhui Tang

专题命中 文生图 :text-to-image(title,abstract);分类 cs.MM

Comments Accepted by ACM MM2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.20495 2024-08-29 eess.IV cs.CV 79%

Enhancing Quantitative Image Synthesis through Pretraining and Resolution Scaling for Bone Mineral Density Estimation from a Plain X-ray Image

Yi Gu, Yoshito Otake, Keisuke Uemura, Masaki Takao, Mazen Soufi, Seiji Okada, Nobuhiko Sugano, Hugues Talbot, Yoshinobu Sato

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments SASHIMI, 2024 (MICCAI workshop). 13 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.18967 2024-08-29 cs.CV 79%

Structural Attention: Rethinking Transformer for Unpaired Medical Image Synthesis

Vu Minh Hieu Phan, Yutong Xie, Bowen Zhang, Yuankai Qi, Zhibin Liao, Antonios Perperidis, Son Lam Phung, Johan W. Verjans, Minh-Son To

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments MICCAI version before camera ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.15261 2024-08-29 cs.HC cs.AI cs.CV cs.IR 79%

Civiverse: A Dataset for Analyzing User Engagement with Open-Source Text-to-Image Models

Maria-Teresa De Rosa Palmini, Laura Wagner, Eva Cetinic

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.13149 2024-08-27 cs.CV 79%

Focus on Neighbors and Know the Whole: Towards Consistent Dense Multiview Text-to-Image Generator for 3D Creation

Bonan Li, Zicheng Zhang, Xingyi Yang, Xinchao Wang

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.01197 2024-08-07 cs.CV 79%

Getting it Right: Improving Spatial Consistency in Text-to-Image Models

Agneet Chatterjee, Gabriela Ben Melech Stan, Estelle Aflalo, Sayak Paul, Dhruba Ghosh, Tejas Gokhale, Ludwig Schmidt, Hannaneh Hajishirzi, Vasudev Lal, Chitta Baral, Yezhou Yang

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

Comments Accepted to ECCV 2024. Project Page : https://spright-t2i.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.11889 2024-08-01 eess.IV cs.CV 79%

Multi-view X-ray Image Synthesis with Multiple Domain Disentanglement from CT Scans

Lixing Tan, Shuang Song, Kangneng Zhou, Chengbo Duan, Lanying Wang, Huayang Ren, Linlin Liu, Wei Zhang, Ruoxiu Xiao

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments 13 pages, 10 figures, ACM MM2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.03411 2024-07-26 cs.CV 79%

Interactive Text-to-Image Retrieval with Large Language Models: A Plug-and-Play Approach

Saehyung Lee, Sangwon Yu, Junsung Park, Jihun Yi, Sungroh Yoon

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

Comments ACL 2024 Oral

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.13277 2024-07-19 eess.IV cs.CV 79%

URCDM: Ultra-Resolution Image Synthesis in Histopathology

Sarah Cechnicka, James Ball, Matthew Baugh, Hadrien Reynaud, Naomi Simmonds, Andrew P. T. Smith, Catherine Horsfield, Candice Roufosse, Bernhard Kainz

专题命中 文生图 :image synthesis(title);diffusion(abstract);分类 cs.CV

Comments arXiv admin note: text overlap with arXiv:2312.01152

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.01261 2024-07-18 cs.CV cs.CY 79%

TIBET: Identifying and Evaluating Biases in Text-to-Image Generative Models

Aditya Chinchure, Pushkar Shukla, Gaurav Bhatt, Kiri Salij, Kartik Hosanagar, Leonid Sigal, Matthew Turk

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

Comments Accepted to ECCV 2024. Code and data available at https://tibet-ai.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.07660 2024-07-11 cs.CV cs.AI 79%

Boosting Medical Image Synthesis via Registration-guided Consistency and Disentanglement Learning

Chuanpu Li, Zeli Chen, Yiwen Zhang, Liming Zhong, Wei Yang

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.05340 2024-07-10 cs.CV eess.IV 79%

Unified Multi-Modal Image Synthesis for Missing Modality Imputation

Yue Zhang, Chengtao Peng, Qiuli Wang, Dan Song, Kaiyan Li, S. Kevin Zhou

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments IEEE TMI accepted final version

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.04940 2024-07-02 cs.CV 79%

Harnessing the Power of MLLMs for Transferable Text-to-Image Person ReID

Wentao Tan, Changxing Ding, Jiayu Jiang, Fei Wang, Yibing Zhan, Dapeng Tao

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.18545 2024-06-28 cs.CV cs.LG 79%

Visual Analysis of Prediction Uncertainty in Neural Networks for Deep Image Synthesis

Soumya Dutta, Faheem Nizar, Ahmad Amaan, Ayan Acharya

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.18400 2024-06-28 cs.MM 79%

Towards Alleviating Text-to-Image Retrieval Hallucination for CLIP in Zero-shot Learning

Hanyao Wang, Yibing Zhan, Liu Liu, Liang Ding, Yan Yang, Jun Yu

专题命中 文生图 :text-to-image(title,abstract);分类 cs.MM

Comments This work has been submitted to the lEEE for possible publication. Copyright may betransferred without notice, after which this version may no longer be accessible

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.02153 2024-06-05 cs.CV 79%

Analyzing the Feature Extractor Networks for Face Image Synthesis

Erdi Sarıtaş, Hazım Kemal Ekenel

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments Accepted at 18th International Conference on Automatic Face and Gesture Recognition (FG) on 1st SD-FGA Workshop 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.00687 2024-06-05 cs.CV 79%

Lay-A-Scene: Personalized 3D Object Arrangement Using Text-to-Image Priors

Ohad Rahamim, Hilit Segev, Idan Achituve, Yuval Atzmon, Yoni Kasten, Gal Chechik

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.02118 2024-05-29 cs.CY cs.AI cs.CV 79%

Position: Towards Implicit Prompt For Text-To-Image Models

Yue Yang, Yuqi Lin, Hong Liu, Wenqi Shao, Runjian Chen, Hailong Shang, Yu Wang, Yu Qiao, Kaipeng Zhang, Ping Luo

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.14768 2024-04-24 cs.CV 79%

Enhancing Prompt Following with Visual Control Through Training-Free Mask-Guided Diffusion

Hongyu Chen, Yiqi Gao, Min Zhou, Peng Wang, Xubin Li, Tiezheng Ge, Bo Zheng

专题命中 文生图 :diffusion(title);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.12541 2024-04-22 cs.CV 79%

GenVideo: One-shot Target-image and Shape Aware Video Editing using T2I Diffusion Models

Sai Sree Harsha, Ambareesh Revanur, Dhwanit Agarwal, Shradha Agrawal

专题命中 文生图 :diffusion(title,abstract);分类 cs.CV

Comments CVPRw 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.16278 2024-04-18 cs.CV 79%

VehicleGAN: Pair-flexible Pose Guided Image Synthesis for Vehicle Re-identification

Baolu Li, Ping Liu, Lan Fu, Jinlong Li, Jianwu Fang, Zhigang Xu, Hongkai Yu

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.14987 2024-04-17 cs.CV 79%

Generative Active Learning for Image Synthesis Personalization

Xulu Zhang, Wengyu Zhang, Xiao-Yong Wei, Jinlin Wu, Zhaoxiang Zhang, Zhen Lei, Qing Li

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.06240 2024-04-10 cs.CV 79%

Hyperparameter-Free Medical Image Synthesis for Sharing Data and Improving Site-Specific Segmentation

Alexander Chebykin, Peter A. N. Bosman, Tanja Alderliesten

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments Accepted at MIDL 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.17275 2024-04-03 cs.CV 79%

One-Shot Structure-Aware Stylized Image Synthesis

Hansam Cho, Jonghyun Lee, Seunggyu Chang, Yonghyun Jeong

专题命中 文生图 :image synthesis(title);diffusion(abstract);分类 cs.CV

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.18493 2024-03-28 cs.CV 79%

VersaT2I: Improving Text-to-Image Models with Versatile Reward

Jianshu Guo, Wenhao Chai, Jie Deng, Hsiang-Wei Huang, Tian Ye, Yichen Xu, Jiawei Zhang, Jenq-Neng Hwang, Gaoang Wang

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.17804 2024-03-27 cs.CV cs.CL 79%

Improving Text-to-Image Consistency via Automatic Prompt Optimization

Oscar Mañas, Pietro Astolfi, Melissa Hall, Candace Ross, Jack Urbanek, Adina Williams, Aishwarya Agrawal, Adriana Romero-Soriano, Michal Drozdzal

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.02109 2024-03-27 cs.CV 79%

ArtAdapter: Text-to-Image Style Transfer using Multi-Level Style Encoder and Explicit Adaptation

Dar-Yen Chen, Hamish Tennent, Ching-Wen Hsu

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.13964 2024-03-26 cs.CV cs.AI 79%

PIA: Your Personalized Image Animator via Plug-and-Play Modules in Text-to-Image Models

Yiming Zhang, Zhening Xing, Yanhong Zeng, Youqing Fang, Kai Chen

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

Comments Project page: https://pi-animator.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏