arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 86585 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 3480 篇

2312.02017 2023-12-05 eess.IV cs.CV physics.med-ph 57%

A multi-channel cycleGAN for CBCT to CT synthesis

Chelsea A. H. Sargeant, Edward G. A. Henderson, Dónal M. McSweeney, Aaron G. Rankin, Denis Page

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments RRRocket_Lollies submission for the Synthesizing computed tomography for radiotherapy (SynthRAD2023) Challenge at MICCAI 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.01300 2023-12-04 cs.CV 57%

Revisiting DETR Pre-training for Object Detection

Yan Ma, Weicong Liang, Bohan Chen, Yiduo Hao, Bojian Hou, Xiangyu Yue, Chao Zhang, Yuhui Yuan

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.18614 2023-12-01 cs.CV 57%

Anatomy and Physiology of Artificial Intelligence in PET Imaging

Tyler J. Bradshaw, Alan B. McMillan

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Journal ref PET Clin; 16(4):471-482 (2021)

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.11452 2023-12-01 cs.CV 57%

ReDirTrans: Latent-to-Latent Translation for Gaze and Head Redirection

Shiwei Jin, Zhen Wang, Lei Wang, Ning Bi, Truong Nguyen

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.15477 2023-11-28 cs.CV 57%

DreamCreature: Crafting Photorealistic Virtual Creatures from Imagination

Kam Woh Ng, Xiatian Zhu, Yi-Zhe Song, Tao Xiang

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments Website: https://github.com/kamwoh/dreamcreature

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.05697 2023-11-28 eess.IV cs.CV 57%

3DGAUnet: 3D generative adversarial networks with a 3D U-Net based generator to achieve the accurate and effective synthesis of clinical tumor image data for pancreatic cancer

Yu Shi, Hannah Tang, Michael Baine, Michael A. Hollingsworth, Huijing Du, Dandan Zheng, Chi Zhang, Hongfeng Yu

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Published on Cancers: Shi, Yu, Hannah Tang, Michael J. Baine, Michael A. Hollingsworth, Huijing Du, Dandan Zheng, Chi Zhang, and Hongfeng Yu. 2023. "3DGAUnet: 3D Generative Adversarial Networks with a 3D U-Net Based Generator to Achieve the Accurate and Effective Synthesis of Clinical Tumor Image Data for Pancreatic Cancer" Cancers 15, no. 23: 5496

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.14240 2023-11-02 cs.MM 57%

Boon: A Neural Search Engine for Cross-Modal Information Retrieval

Yan Gong, Georgina Cosma

专题命中 文生图 :text-to-image(abstract);分类 cs.MM

Journal ref MMIR '23: Proceedings of the 1st International Workshop on Deep Multimodal Learning for Information Retrieval (2023)

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.19297 2023-10-31 cs.LG cs.CV cs.CY 57%

On Measuring Fairness in Generative Models

Christopher T. H. Teo, Milad Abdollahzadeh, Ngai-Man Cheung

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments Accepted in NeurIPS23

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.18894 2023-10-31 cs.CV 57%

Emergence of Shape Bias in Convolutional Neural Networks through Activation Sparsity

Tianqin Li, Ziqi Wen, Yangfan Li, Tai Sing Lee

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Published as NeurIPS 2023 (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.00157 2023-10-27 cs.CV 57%

Ponder: Point Cloud Pre-training via Neural Rendering

Di Huang, Sida Peng, Tong He, Honghui Yang, Xiaowei Zhou, Wanli Ouyang

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Project page: https://dihuang.me/ponder/

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.16255 2023-10-26 cs.CV 57%

UAV-Sim: NeRF-based Synthetic Data Generation for UAV-based Perception

Christopher Maxey, Jaehoon Choi, Hyungtae Lee, Dinesh Manocha, Heesung Kwon

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Video Link: https://www.youtube.com/watch?v=ucPzbPLqqpI

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.02397 2023-10-18 cs.CV 57%

A Scene-Text Synthesis Engine Achieved Through Learning from Decomposed Real-World Data

Zhengmi Tang, Tomo Miyazaki, Shinichiro Omachi

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.08921 2023-10-16 cs.CV 57%

Feature Proliferation -- the "Cancer" in StyleGAN and its Treatments

Shuang Song, Yuanbang Liang, Jing Wu, Yu-Kun Lai, Yipeng Qin

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Accepted at ICCV 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.06904 2023-10-12 cs.CV 57%

Mitigating stereotypical biases in text to image generative systems

Piero Esposito, Parmida Atighehchian, Anastasis Germanidis, Deepti Ghadiyaram

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments 4 figures, 8 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.20062 2023-10-06 cs.CV 57%

Chatting Makes Perfect: Chat-based Image Retrieval

Matan Levy, Rami Ben-Ari, Nir Darshan, Dani Lischinski

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments Camera Ready version for NeurIPS 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.02936 2023-10-06 cs.LG cs.CR cs.CV 57%

Private GANs, Revisited

Alex Bie, Gautam Kamath, Guojun Zhang

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments 28 pages; revisions and new experiments from TMLR camera-ready + code release at https://github.com/alexbie98/dpgan-revisit

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.07623 2023-09-15 cs.CV 57%

SwitchGPT: Adapting Large Language Models for Non-Text Outputs

Xinyu Wang, Bohan Zhuang, Qi Wu

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.15955 2023-09-08 cs.CV 57%

Understanding Prompt Tuning for V-L Models Through the Lens of Neural Collapse

Didi Zhu, Zexi Li, Min Zhang, Junkun Yuan, Yunfeng Shao, Jiashuo Liu, Kun Kuang, Yinchuan Li, Chao Wu

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.13853 2023-08-29 cs.CV 57%

Beyond One-to-One: Rethinking the Referring Image Segmentation

Yutao Hu, Qixiong Wang, Wenqi Shao, Enze Xie, Zhenguo Li, Jungong Han, Ping Luo

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments ICCV 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.04308 2023-08-23 cs.CV 57%

Enhancing Modality-Agnostic Representations via Meta-Learning for Brain Tumor Segmentation

Aishik Konwer, Xiaoling Hu, Joseph Bae, Xuan Xu, Chao Chen, Prateek Prasanna

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Accepted in ICCV 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.14797 2023-08-10 cs.CV cs.LG 57%

3D-Aware Video Generation

Sherwin Bahmani, Jeong Joon Park, Despoina Paschalidou, Hao Tang, Gordon Wetzstein, Leonidas Guibas, Luc Van Gool, Radu Timofte

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments TMLR 2023; Project page: https://sherwinbahmani.github.io/3dvidgen

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.04052 2023-08-09 cs.LG cs.CL cs.CV 57%

The Five-Dollar Model: Generating Game Maps and Sprites from Sentence Embeddings

Timothy Merino, Roman Negri, Dipika Rajesh, M Charity, Julian Togelius

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments to be published in AIIDE 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.16143 2023-08-02 eess.IV cs.CV 57%

Structure-Preserving Synthesis: MaskGAN for Unpaired MR-CT Translation

Minh Hieu Phan, Zhibin Liao, Johan W. Verjans, Minh-Son To

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Accepted to MICCAI 2023

Journal ref MICCAI 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.16021 2023-08-01 eess.IV cs.CV 57%

LOTUS: Learning to Optimize Task-based US representations

Yordanka Velikova, Mohammad Farid Azampour, Walter Simson, Vanessa Gonzalez Duque, Nassir Navab

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Accepted at International Conference on Medical Image Computing and Computer Assisted Intervention, MICCAI 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.15860 2023-08-01 cs.CV cs.CR 57%

What can Discriminator do? Towards Box-free Ownership Verification of Generative Adversarial Network

Ziheng Huang, Boheng Li, Yan Cai, Run Wang, Shangwei Guo, Liming Fang, Jing Chen, Lina Wang

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Accepted to ICCV 2023. The first two authors contributed equally to this work

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.06350 2023-07-31 cs.CV 57%

VITR: Augmenting Vision Transformers with Relation-Focused Learning for Cross-Modal Information Retrieval

Yan Gong, Georgina Cosma, Axel Finke

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.14244 2023-07-27 cs.MM 57%

Neural-based Cross-modal Search and Retrieval of Artwork

Yan Gong, Georgina Cosma, Axel Finke

专题命中 文生图 :text-to-image(abstract);分类 cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.12914 2023-07-26 cs.CV cs.AI 57%

Towards a Visual-Language Foundation Model for Computational Pathology

Ming Y. Lu, Bowen Chen, Drew F. K. Williamson, Richard J. Chen, Ivy Liang, Tong Ding, Guillaume Jaume, Igor Odintsov, Andrew Zhang, Long Phi Le, Georg Gerber, Anil V Parwani, Faisal Mahmood

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.10776 2023-07-21 cs.CV 57%

Urban Radiance Field Representation with Deformable Neural Mesh Primitives

Fan Lu, Yan Xu, Guang Chen, Hongsheng Li, Kwan-Yee Lin, Changjun Jiang

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Accepted to ICCV2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.07341 2023-07-20 cs.IR cs.CV 57%

PiTL: Cross-modal Retrieval with Weakly-supervised Vision-language Pre-training via Prompting

Zixin Guo, Tzu-Jui Julius Wang, Selen Pehlivan, Abduljalil Radman, Jorma Laaksonen

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Journal ref SIGIR, 2023, 2261-2265

详情

展开后加载摘要…

URL PDF HTML 收藏