Navigating the Synthetic Realm: Harnessing Diffusion-based Models for Laparoscopic Text-to-Image Generation
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image generation(title);分类 cs.CV
视觉与机器人
图像生成、文生图、图像编辑、扩散模型和可控生成。
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image generation(title);分类 cs.CV
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image synthesis(title);分类 cs.CV
Comments Accepted by ICCV 2023. Code is available at: https://github.com/showlab/BoxDiff
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image generation(title);分类 cs.CV
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image generation(title);分类 cs.CV
SMPL-GPTexture:利用文本到图像生成模型进行双视角3D人体纹理估计
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);diffusion(abstract);inpainting(abstract)
AI总结 本文提出SMPL-GPTexture方法,通过文本提示生成双视角图像,结合人体网格恢复模型和反向光栅化技术,生成高精度纹理映射,解决3D人体纹理生成中的隐私和数据获取难题。
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);diffusion(abstract);image editing(abstract)
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image generation(abstract);image synthesis(abstract)
Comments Accepted to WACV 2025
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image generation(abstract);image synthesis(abstract)
Comments Carmera-ready version. To appear in ACM MM 2023. Code will be released at: https://github.com/sf-zhai/BadT2I
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);diffusion(abstract);image editing(abstract)
HADIS:混合自适应扩散模型服务用于高效的文本到图像生成
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image generation(title)
AI总结 HADIS通过混合自适应架构优化扩散模型服务,提升响应质量和降低延迟违规率
Comments 15 pages, 12 figures
用于安全文本到图像生成的内省注意力调制
机构 * The University of Melbourne(墨尔本大学) ; Lancaster University(兰卡斯特大学)
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);diffusion(abstract);image synthesis(abstract)
AI总结 研究基于流的文本到图像模型易产生不安全内容的问题,提出通过推理时内省调节注意力动态实现安全的方法,该方法在保持或提升质量的同时显著提高安全分数,为安全图像生成提供新途径。
Comments Accepted at ECCV 2026. 20 pages, 7 figures. Project page: https://basim-azam.github.io/iam/
编辑:用于文本到图像扩散模型的有效且可解释的提示倒置
机构 * University of Massachusetts, Amherst(马萨诸塞大学阿姆赫斯特分校) ; Georgia Institute of Technology(佐治亚理工学院) ; Rutgers University(罗格斯大学) ; University of Utah(犹他大学) ; Dolby Laboratories(杜比实验室)
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image generation(abstract);image synthesis(abstract)
AI总结 本文提出\sys技术,通过预训练模型初始化、潜在空间反向工程和嵌入到文本转换,提升文本到图像扩散模型的提示倒置效果,实现更高的图像相似性和可解释性。
机构 * Departments of Convergence Medicine, University of Ulsan College of Medicine, Asan Medical Center(融合医学部门、首尔大学医学院、Asan医院) ; Department of Otorhinolaryngology-Head and Neck Surgery, University of Ulsan College of Medicine, Asan Medical Center(耳鼻喉科-头颈外科部门、首尔大学医学院、Asan医院)
专题命中 文生图 :text-to-image(title,abstract);image synthesis(title,abstract);image generation(abstract);diffusion(abstract)
机构 * BRAC University(布拉克大学)
专题命中 文生图 :text-to-image(title,abstract);image synthesis(title,abstract);image generation(abstract);diffusion(abstract)
Comments 19 pages, 36 figures
机构 * Department of Computer Science, SCMNS School, Morgan State University, Baltimore, Maryland 21251, USA ; Electrical \& Computer Engineering Dept., School of Engineering, Morgan State University, Baltimore, Maryland 21251, USA
专题命中 文生图 :diffusion(title,abstract);image synthesis(title,abstract);image generation(abstract);text-to-image(abstract)
Journal ref Conference and Labs of the Evaluation Forum (CLEF) 2024
机构 * Data Science and AI Laboratory, ECE, Seoul National University(数据科学与人工智能实验室、电子与计算机工程系、首尔国立大学) ; AIIS, ASRI, INMC, ISRC, and Interdisciplinary Program in AI, Seoul National University(人工智能研究所、人工智能研究室、智能网络与计算中心、信息科学与工程系以及人工智能交叉学科项目、首尔国立大学)
专题命中 文生图 :text-to-image(title,abstract);inpainting(title,abstract);image generation(abstract);image editing(abstract)
Comments CVPR 2025
机构 * The Hong Kong Polytechnic University(香港理工大学) ; University of Washington(华盛顿大学) ; Xi’an Jiaotong University(西安交通大学)
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);diffusion(abstract);image synthesis(abstract)
Comments 10 pages, 4 figures
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);diffusion(abstract);image synthesis(abstract)
Comments Project page: https://github.com/KaiyueSun98/T2I-Personalization-with-AR
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);diffusion(abstract);image editing(abstract)
Comments 75 pages, 73 figures, Evaluation scripts: https://github.com/jylei16/Imagine-e
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);diffusion(abstract);image synthesis(abstract)
Comments Accepted by ICLR2025
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);diffusion(abstract);image synthesis(abstract)
Comments Under review
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);diffusion(abstract);image synthesis(abstract)
Comments COLM 2024, Project Url: https://zeyofu.github.io/CommonsenseT2I/
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);diffusion(abstract);image synthesis(abstract)
Comments Accepted by ECCV 2024, code and dataset available in https://github.com/open-mmlab/AnyControl
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);diffusion(abstract);image editing(abstract)
Comments This paper has been accepted for oral presentation at the IJCAI 2024 Workshop on Trustworthy Interactive Decision-Making with Foundation Models
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image generation(abstract);image synthesis(abstract)
Comments Corresponding to "Würstchen v2"
Journal ref The Twelfth International Conference on Learning Representations (ICLR), 2024
专题命中 文生图 :diffusion(title,abstract);image synthesis(title,abstract);image generation(abstract);image editing(abstract)
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);inpainting(abstract);image synthesis(abstract)
Comments Accepted by CVPR2024. Codes and models are released at https://github.com/ewrfcas/LeftRefill, Project page: https://ewrfcas.github.io/LeftRefill
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image generation(abstract);image synthesis(abstract)
专题命中 文生图 :text-to-image(title,abstract);image synthesis(title,abstract);image generation(abstract);diffusion(abstract)
Comments Published in the Journal of Artificial Intelligence Research (JAIR)
Journal ref Journal of Artificial Intelligence Research (JAIR), Vol. 78 (2023)
专题命中 文生图 :text-to-image(title,abstract);image editing(title,abstract);image generation(abstract);diffusion(abstract)
Comments Full paper is accepted by WACV2024; Best paper runner-up of AI4CC@CVPR 2023