arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 86585 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 3480 篇

2207.02250 2022-07-07 cs.CV eess.IV 57%

Array Camera Image Fusion using Physics-Aware Transformers

Qian Huang, Minghao Hu, David Jones Brady

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.11404 2022-06-24 cs.CV cs.AI cs.LG 57%

The ArtBench Dataset: Benchmarking Generative Models with Artworks

Peiyuan Liao, Xiuyu Li, Xihui Liu, Kurt Keutzer

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments The first two authors contributed equally to this work. The code and data are available at https://github.com/liaopeiyuan/artbench

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.04503 2022-06-10 cs.CV cs.AI 57%

cycle text2face: cycle text-to-face gan via transformers

Faezeh Gholamrezaie, Mohammad Manthouri

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.15868 2022-06-01 cs.CV cs.CL cs.LG 57%

CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers

Wenyi Hong, Ming Ding, Wendi Zheng, Xinghan Liu, Jie Tang

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.13294 2022-05-27 cs.CV eess.IV eess.SP 57%

Analytical Interpretation of Latent Codes in InfoGAN with SAR Images

Zhenpeng Feng, Milos Dakovic, Hongbing Ji, Mingzhe Zhu, Ljubisa Stankovic

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments 13 pages, 14 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.07441 2022-05-23 cs.CV cs.CL cs.IR 57%

COTS: Collaborative Two-Stream Vision-Language Pre-Training Model for Cross-Modal Retrieval

Haoyu Lu, Nanyi Fei, Yuqi Huo, Yizhao Gao, Zhiwu Lu, Ji-Rong Wen

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments Accepted by CVPR2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.08101 2022-05-17 cs.CV cs.IR 57%

ARTEMIS: Attention-based Retrieval with Text-Explicit Matching and Implicit Similarity

Ginger Delmas, Rafael Sampaio de Rezende, Gabriela Csurka, Diane Larlus

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments Published in ICLR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.06126 2022-05-13 cs.CL cs.CV cs.LG 57%

One Model, Multiple Modalities: A Sparsely Activated Approach for Text, Sound, Image, Video and Code

Yong Dai, Duyu Tang, Liangxin Liu, Minghuan Tan, Cong Zhou, Jingquan Wang, Zhangyin Feng, Fan Zhang, Xueyu Hu, Shuming Shi

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.05922 2022-05-13 cs.CV 57%

Ray Priors through Reprojection: Improving Neural Radiance Fields for Novel View Extrapolation

Jian Zhang, Yuanqing Zhang, Huan Fu, Xiaowei Zhou, Bowen Cai, Jinchi Huang, Rongfei Jia, Binqiang Zhao, Xing Tang

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.00273 2022-05-06 cs.LG cs.CV 57%

StyleGAN-XL: Scaling StyleGAN to Large Diverse Datasets

Axel Sauer, Katja Schwarz, Andreas Geiger

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments To appear in SIGGRAPH 2022. Project Page: https://sites.google.com/view/stylegan-xl/

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.01879 2022-05-05 cs.CV cs.AI 57%

Learning Two-Stream CNN for Multi-Modal Age-related Macular Degeneration Categorization

Weisen Wang, Xirong Li, Zhiyan Xu, Weihong Yu, Jianchun Zhao, Dayong Ding, Youxin Chen

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Accepted by IEEE Journal of Biomedical and Health Informatics (J-BHI)

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.12797 2022-04-28 cs.GR 57%

Towards Quantum Ray Tracing

Luís Paulo Santos, Thomas Bashford-Rogers, João Barbosa, Paul Navrátil

专题命中 文生图 :image synthesis(abstract);分类 cs.GR

Comments 27 pages, 15 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.12007 2022-04-28 eess.IV cs.CV physics.med-ph 57%

Assessing the ability of generative adversarial networks to learn canonical medical image statistics

Varun A. Kelkar, Dimitrios S. Gotsis, Frank J. Brooks, Prabhat KC, Kyle J. Myers, Rongping Zeng, Mark A. Anastasio

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.08075 2022-04-28 cs.CL cs.CV 57%

Things not Written in Text: Exploring Spatial Commonsense from Visual Signals

Xiao Liu, Da Yin, Yansong Feng, Dongyan Zhao

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Accepted by ACL 2022 main conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.05864 2022-04-26 cs.CV cs.AI 57%

Human Silhouette and Skeleton Video Synthesis through Wi-Fi signals

Danilo Avola, Marco Cascio, Luigi Cinque, Alessio Fagioli, Gian Luca Foresti

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Journal ref International Journal of Neural Systems, 2022, 2250015

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.09268 2022-04-21 cs.LG cs.CL cs.CV cs.IR 57%

Uncertainty-based Cross-Modal Retrieval with Probabilistic Representations

Leila Pishdad, Ran Zhang, Konstantinos G. Derpanis, Allan Jepson, Afsaneh Fazly

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments 13 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.08339 2022-04-19 cs.CV 57%

Migrating Face Swap to Mobile Devices: A lightweight Framework and A Supervised Training Solution

Haiming Yu, Hao Zhu, Xiangju Lu, Junhui Liu

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Accepted to IEEE International Conference on Multimedia and Expo 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.03547 2022-04-08 eess.IV cs.CV physics.med-ph 57%

Evaluating Procedures for Establishing Generative Adversarial Network-based Stochastic Image Models in Medical Imaging

Varun A. Kelkar, Dimitrios S. Gotsis, Frank J. Brooks, Kyle J. Myers, Prabhat KC, Rongping Zeng, Mark A. Anastasio

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Published in SPIE Medical Imaging 2022: Image Perception, Observer Performance, and Technology Assessment

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.01960 2022-04-06 cs.CV cs.AI stat.ML 57%

FaceSigns: Semi-Fragile Neural Watermarks for Media Authentication and Countering Deepfakes

Paarth Neekhara, Shehzeen Hussain, Xinqiao Zhang, Ke Huang, Julian McAuley, Farinaz Koushanfar

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments 13 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.03959 2022-04-01 cs.LG cs.CV 57%

Generative Flows with Invertible Attentions

Rhea Sanjay Sukthanker, Zhiwu Huang, Suryansh Kumar, Radu Timofte, Luc Van Gool

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Accepted to CVPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.15006 2022-03-30 cs.CV 57%

TL-GAN: Improving Traffic Light Recognition via Data Synthesis for Autonomous Driving

Danfeng Wang, Xin Ma, Xiaodong Yang

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.09824 2022-03-21 cs.CV cs.LG eess.AS 57%

Cross-Modal Perceptionist: Can Face Geometry be Gleaned from Voices?

Cho-Ying Wu, Chin-Cheng Hsu, Ulrich Neumann

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Accepted to CVPR 2022. Project page: https://choyingw.github.io/works/Voice2Mesh/index.html. This version supersedes arXiv:2104.10299

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.09771 2022-03-21 cs.CV 57%

Beyond a Video Frame Interpolator: A Space Decoupled Learning Approach to Continuous Image Transition

Tao Yang, Peiran Ren, Xuansong Xie, Xiansheng Hua, Lei Zhang

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.03638 2022-03-09 eess.IV cs.AI cs.CV cs.LG 57%

Unsupervised Image Registration Towards Enhancing Performance and Explainability in Cardiac And Brain Image Analysis

Chengjia Wang, Guang Yang, Giorgos Papanastasiou

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments 38 pages, 7 figures, will be published in Sensors journal by MDPI

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.01716 2022-03-04 cs.CR cs.CV cs.LG eess.IV 57%

Detecting High-Quality GAN-Generated Face Images using Neural Networks

Ehsan Nowroozi, Mauro Conti, Yassine Mekdad

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments 16 Pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.12752 2022-02-28 cs.CV 57%

Synthesizing Photorealistic Images with Deep Generative Learning

Chuanxia Zheng

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments PhD thesis

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.12211 2022-02-25 cs.CV 57%

Self-Distilled StyleGAN: Towards Generation from Internet Photos

Ron Mokady, Michal Yarom, Omer Tov, Oran Lang, Daniel Cohen-Or, Tali Dekel, Michal Irani, Inbar Mosseri

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.10687 2022-02-23 cs.CV 57%

Cut and Continuous Paste towards Real-time Deep Fall Detection

Sunhee Hwang, Minsong Ki, Seung-Hyun Lee, Sanghoon Park, Byoung-Ki Jeon

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Accepted to ICASSP 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.09205 2022-02-04 cs.CV cs.CR 57%

Adversarial Rain Attack and Defensive Deraining for DNN Perception

Liming Zhai, Felix Juefei-Xu, Qing Guo, Xiaofei Xie, Lei Ma, Wei Feng, Shengchao Qin, Yang Liu

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2201.09152 2022-01-25 cs.CV cs.LG eess.IV 57%

Generative Adversarial Network Applications in Creating a Meta-Universe

Soheyla Amirian, Thiab R. Taha, Khaled Rasheed, Hamid R. Arabnia

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments Computational Science and Computational Intelligence; 2021 International Conference on IEEE CPS (IEEE XPLORE, Scopus), IEEE, 2021

详情

展开后加载摘要…

URL PDF HTML 收藏