arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 86539 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 3480 篇

1905.05947 2019-05-16 cs.CV 79%

Joint haze image synthesis and dehazing with mmd-vae losses

Zongliang Li, Chi Zhang, Gaofeng Meng, Yuehu Liu

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments Preprinted version on arxiv, May-05-2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1812.01002 2019-04-29 cs.CV 79%

Disentangling Latent Hands for Image Synthesis and Pose Estimation

Linlin Yang, Angela Yao

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments CVPR 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.05118 2019-04-11 cs.CV 79%

Text Guided Person Image Synthesis

Xingran Zhou, Siyu Huang, Bin Li, Yingming Li, Jiachen Li, Zhongfei Zhang

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments To appear at CVPR 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1810.10850 2019-04-09 cs.CV 79%

An Adversarial Learning Approach to Medical Image Synthesis for Lesion Detection

Liyan Sun, Jiexiang Wang, Yue Huang, Xinghao Ding, Hayit Greenspan, John Paisley

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments 10 pages, 13 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.01782 2019-04-04 cs.CV 79%

Conditional Adversarial Generative Flow for Controllable Image Synthesis

Rui Liu, Yu Liu, Xinyu Gong, Xiaogang Wang, Hongsheng Li

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments Accepted by CVPR 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1812.05840 2019-04-03 cs.CV 79%

Spatial Fusion GAN for Image Synthesis

Fangneng Zhan, Hongyuan Zhu, Shijian Lu

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments Accepted to CVPR 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1901.06219 2019-03-11 cs.CV 79%

Red blood cell image generation for data augmentation using Conditional Generative Adversarial Networks

Oleksandr Bailo, DongShik Ham, Young Min Shin

专题命中 文生图 :image generation(title);image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1903.02225 2019-03-07 cs.CV eess.IV 79%

DepthwiseGANs: Fast Training Generative Adversarial Networks for Realistic Image Synthesis

Mkhuseli Ngxande, Jules-Raymond Tapamo, Michael Burke

专题命中 文生图 :image synthesis(title);text-to-image(abstract);分类 cs.CV

Comments 6 pages, 8 figures, To appear in the Proceedings of Southern African Universities Power EngineeringConference/Robotics and Mechatronics/Pattern Recognition Association of South Africa(SAUPEC/RobMech/PRASA), January 20-30 2019, Bloemfotein, South Africa

详情

展开后加载摘要…

URL PDF HTML 收藏
1901.06322 2019-01-21 cs.CV cs.LG 79%

Learning Spatial Pyramid Attentive Pooling in Image Synthesis and Image-to-Image Translation

Wei Sun, Tianfu Wu

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments 12 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
1810.11610 2019-01-14 cs.CV 79%

Soft-Gated Warping-GAN for Pose-Guided Person Image Synthesis

Haoye Dong, Xiaodan Liang, Ke Gong, Hanjiang Lai, Jia Zhu, Jian Yin

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments 17 pages, 14 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1812.08985 2018-12-24 cs.LG cs.CV stat.ML 79%

Non-Adversarial Image Synthesis with Generative Latent Nearest Neighbors

Yedid Hoshen, Jitendra Malik

专题命中 文生图 :image synthesis(title);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1812.04652 2018-12-13 cs.CV 79%

Evaluating the Impact of Intensity Normalization on MR Image Synthesis

Jacob C. Reinhold, Blake E. Dewey, Aaron Carass, Jerry L. Prince

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments SPIE Medical Imaging 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1806.03796 2018-11-21 cs.CV cs.LG cs.NE stat.ML 79%

Generative Adversarial Network Architectures For Image Synthesis Using Capsule Networks

Yash Upadhyay, Paul Schrater

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments Pre-print

详情

展开后加载摘要…

URL PDF HTML 收藏
1807.03021 2018-09-27 cs.CV cs.AI 79%

Verisimilar Image Synthesis for Accurate Detection and Recognition of Texts in Scenes

Fangneng Zhan, Shijian Lu, Chuhui Xue

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments 14 pages, ECCV2018, datasets: https://github.com/fnzhan/Verisimilar-Image-Synthesis-for-Accurate-Detection-and-Recognition-of-Texts-in-Scenes

详情

展开后加载摘要…

URL PDF HTML 收藏
1808.04538 2018-08-15 cs.LG cs.CL cs.CV stat.ML 79%

Text-to-Image-to-Text Translation using Cycle Consistent Adversarial Networks

Satya Krishna Gorti, Jeremy Ma

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1803.06629 2018-08-01 cs.CV 79%

Cross-modality image synthesis from unpaired data using CycleGAN: Effects of gradient consistency loss and training data size

Yuta Hiasa, Yoshito Otake, Masaki Takao, Takumi Matsuoka, Kazuma Takashima, Jerry L. Prince, Nobuhiko Sugano, Yoshinobu Sato

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments 10 pages, 7 figures, MICCAI 2018 Workshop on Simulation and Synthesis in Medical Imaging

详情

展开后加载摘要…

URL PDF HTML 收藏
1806.11475 2018-07-02 cs.CV 79%

SynNet: Structure-Preserving Fully Convolutional Networks for Medical Image Synthesis

Deepa Gunashekar, Sailesh Conjeti, Abhijit Guha Roy, Nassir Navab, Kuangyu Shi

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments 13 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1805.01972 2018-05-08 cs.CV 79%

Fast-converging Conditional Generative Adversarial Networks for Image Synthesis

Chengcheng Li, Zi Wang, Hairong Qi

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments Accepted by ICIP 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1803.03396 2018-03-30 cs.CV 79%

Cross-View Image Synthesis using Conditional GANs

Krishna Regmi, Ali Borji

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments Accepted at CVPR 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1712.07329 2017-12-21 cs.CV 79%

On the Diversity of Realistic Image Synthesis

Zichen Yang, Haifeng Liu, Deng Cai

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

Comments 10 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1708.04692 2017-09-13 cs.CV cs.LG stat.ML 79%

GANs for Biological Image Synthesis

Anton Osokin, Anatole Chessel, Rafael E. Carazo Salas, Federico Vaggi

专题命中 文生图 :image synthesis(title);image generation(abstract);分类 cs.CV

Comments The paper appearing at the International Conference on Computer Vision (ICCV) 2017 + its supplementary materials

详情

展开后加载摘要…

URL PDF HTML 收藏
1612.03242 2017-08-08 cs.CV cs.AI stat.ML 79%

StackGAN: Text to Photo-realistic Image Synthesis with Stacked Generative Adversarial Networks

Han Zhang, Tao Xu, Hongsheng Li, Shaoting Zhang, Xiaogang Wang, Xiaolei Huang, Dimitris Metaxas

专题命中 文生图 :image synthesis(title);text-to-image(abstract);分类 cs.CV

Comments ICCV 2017 Oral Presentation

详情

展开后加载摘要…

URL PDF HTML 收藏
1610.09585 2017-07-24 stat.ML cs.CV 79%

Conditional Image Synthesis With Auxiliary Classifier GANs

Augustus Odena, Christopher Olah, Jonathon Shlens

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1702.01540 2017-02-07 cs.GR 79%

Generalized 3D Voxel Image Synthesis Architecture for Volumetric Spatial Visualization

Anas M. Al-Oraiqat, E. A. Bashkov, S. A. Zori, Aladdein M. Amro

专题命中 文生图 :image synthesis(title,abstract);分类 cs.GR

Comments 9 pages, 4 Figures, International Publisher for Advanced Scientific Journals 2013

详情

展开后加载摘要…

URL PDF HTML 收藏
1609.07370 2016-09-26 cs.CV 79%

Example-Based Image Synthesis via Randomized Patch-Matching

Yi Ren, Yaniv Romano, Michael Elad

专题命中 文生图 :image synthesis(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.11061 2026-05-13 cs.CV cs.MM 79%

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer

HiDream-O1-Image:一种原生统一的图像生成基础模型,具有像素级统一的Transformer

Qi Cai, Jingwen Chen, Chengmin Gao, Zijian Gong, Yehao Li, Yingwei Pan, Yi Peng, Zhaofan Qiu, Kai Yu, Yiheng Zhang, Hao Ai, Siying Bai, Yang Chen, Zhihui Chen, Fengbin Gao, Ying Guo, Dong Li, Zhen Shen, Leilei Shi, Jing Wang, Siyu Wang, Yimeng Wang, Rui Zheng, Ting Yao, Tao Mei

专题命中 文生图 :image generation(abstract);text-to-image(abstract);diffusion(abstract);分类 cs.CV、cs.MM

AI总结 本文提出HiDream-O1-Image,通过像素空间扩散Transformer实现多模态输入的结构统一,无需外部VAE或离散文本编码器,提升生成和编辑任务的性能,实验表明其在多种生成任务中表现优异。

Comments Source codes and models are available at Github: https://github.com/HiDream-ai/HiDream-O1-Image and Huggingface: https://huggingface.co/HiDream-ai/HiDream-O1-Image

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10715 2025-10-14 cs.GR cs.CV 79%

VLM-Guided Adaptive Negative Prompting for Creative Generation

Shelly Golan, Yotam Nitzan, Zongze Wu, Or Patashnik

机构 * Adobe Research(Adobe研究院) Tel Aviv University(特拉维夫大学)

专题命中 文生图 :image generation(abstract);text-to-image(abstract);diffusion(abstract);分类 cs.CV、cs.GR

Comments Project page at: https://shelley-golan.github.io/VLM-Guided-Creative-Generation/

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.20288 2025-05-27 cs.CV cs.MM 79%

Hierarchical Masked Autoregressive Models with Low-Resolution Token Pivots

Guangting Zheng, Yehao Li, Yingwei Pan, Jiajun Deng, Ting Yao, Yanyong Zhang, Tao Mei

专题命中 文生图 :image generation(abstract);text-to-image(abstract);diffusion(abstract);分类 cs.CV、cs.MM

Comments ICML 2025. Source code is available at https://github.com/HiDream-ai/himar

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.15461 2024-12-24 cs.CV cs.MM 79%

Hand1000: Generating Realistic Hands from Text with Only 1,000 Images

Haozhuo Zhang, Bin Zhu, Yu Cao, Yanbin Hao

专题命中 文生图 :image generation(abstract);text-to-image(abstract);diffusion(abstract);分类 cs.CV、cs.MM

Comments Accepted by AAAI 2025. Project page https://haozhuo-zhang.github.io/Hand1000-project-page/

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.02899 2024-04-23 cs.CV cs.GR 79%

MatAtlas: Text-driven Consistent Geometry Texturing and Material Assignment

Duygu Ceylan, Valentin Deschaintre, Thibault Groueix, Rosalie Martin, Chun-Hao Huang, Romain Rouffet, Vladimir Kim, Gaëtan Lassagne

专题命中 文生图 :image generation(abstract);text-to-image(abstract);diffusion(abstract);分类 cs.CV、cs.GR

详情

展开后加载摘要…

URL PDF HTML 收藏