NeuroPrompts: An Adaptive Framework to Optimize Prompts for Text-to-Image Generation
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);diffusion(abstract)
Comments Accepted to EACL 2024 System Demonstration Track
视觉与机器人
图像生成、文生图、图像编辑、扩散模型和可控生成。
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);diffusion(abstract)
Comments Accepted to EACL 2024 System Demonstration Track
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV、cs.GR、cs.MM
Comments Website: https://jiuntian.github.io/interactdiffusion. Accepted at CVPR2024
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);diffusion(abstract)
专题命中 文生图 :text-to-image(title,abstract);image synthesis(title,abstract);image generation(abstract)
BAFIS:评估现代文本到图像模型中的职业偏见与人类偏好的数据集与框架
机构 * RheinMain University of Applied Sciences(莱茵美因应用科学大学)
专题命中 文生图 :text-to-image(title,abstract);diffusion(abstract,abstract_cn);image generation(abstract);image synthesis(abstract)
AI总结 本研究提出BAFIS平台和包含21,140张多语言提示生成图像的数据集,评估五种文本到图像模型在职业生成中的性别和种族偏见,结合人类偏好反馈,发现系统性偏见并强调纳入人类偏好的必要性。
Comments Accepted at the IEEE Winter Conference on Applications of Computer Vision, WACV 2026
专题命中 文生图 :text-to-image(title,abstract);image synthesis(title);image generation(abstract,comments);分类 cs.CV
Comments text-to-image generation, automatic prompt, DPO, Counterfactual
机构 * NextStep-Team(NextStep团队)
专题命中 文生图 :image generation(title,abstract);text-to-image(abstract);diffusion(abstract);image editing(abstract)
Comments Code: https://github.com/stepfun-ai/NextStep-1
ANCHOR:基于大语言模型的文本到图像合成中的主体条件化
机构 * Optum AI ; The Pennsylvania State University(宾夕法尼亚州立大学)
专题命中 文生图 :text-to-image(title,abstract);image synthesis(title,abstract);分类 cs.CV、cs.MM
AI总结 ANCHOR通过大规模抽象式标题数据集研究文本到图像合成中多主体理解与上下文推理的缺陷,提出基于大语言模型的主体感知微调方法,提升图像-标题一致性与人类偏好对齐。
Comments Accepted to The 64th Annual Meeting of the Association for Computational Linguistics (ACL) 2026
通过对比噪声优化实现文本到图像生成的多样性
专题命中 文生图 :text-to-image(title,abstract);image generation(title);diffusion(abstract);分类 cs.CV、cs.GR
AI总结 本文提出对比噪声优化方法,通过调整初始噪声提升文本到图像生成的多样性,同时保持图像质量,实验表明其在质量与多样性之间取得平衡。
Comments Accepted to ICLR 2026
专题命中 文生图 :text-to-image(title,abstract);image generation(title);diffusion(abstract);分类 cs.CV、cs.GR
Journal ref IUI '25: Proceedings of the 30th International Conference on Intelligent User Interfaces, Pages 1381 - 1397, 2025
机构 * Tel-Aviv University(特拉维夫大学) ; Bar-Ilan University(巴伊兰大学)
专题命中 文生图 :text-to-image(title,abstract);diffusion(title);image generation(abstract);分类 cs.CV、cs.GR
Comments Pre-print
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);分类 cs.CV、cs.MM
Comments CVPR 2025 Camera-ready. Project page: https://silmm.github.io/
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);分类 cs.CV、cs.MM
Comments ICLR 2025 Camera-ready
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV、cs.MM
Journal ref International Conference on Learning Representations (ICLR 2025)
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);分类 cs.CV、cs.GR
Comments preprint
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV、cs.MM
Comments ECCV 2024
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV、cs.GR
Comments NeurIPS 2024
专题命中 文生图 :text-to-image(title,abstract);image synthesis(title);diffusion(abstract);分类 cs.CV、cs.MM
Comments Accepted by ACM Multimedia 2024. The dataset and code can be found at https://github.com/achernarwang/LiVO
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);分类 cs.CV、cs.GR
Comments Project website: https://comfygen-paper.github.io/
专题命中 文生图 :text-to-image(title,abstract);diffusion(title);image generation(abstract);分类 cs.CV、cs.GR
Comments Accepted to SIGGRAPH 2024. Project page is available at https://omriavrahami.com/the-chosen-one/
专题命中 文生图 :text-to-image(title,abstract);image generation(title);diffusion(abstract);分类 cs.CV、cs.GR
Comments Project page is at https://make-it-count-paper.github.io/
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);分类 cs.CV、cs.GR
Comments Accepted to journal track of SIGGRAPH 2024 (TOG). Project page is at https://consistory-paper.github.io
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);分类 cs.CV、cs.MM
专题命中 文生图 :text-to-image(title,abstract);image generation(title);diffusion(abstract);分类 cs.CV、cs.GR
Comments Project page: https://omer11a.github.io/bounded-attention/
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);分类 cs.CV、cs.MM
Comments CVPR 2024; project page: https://dpt-t2i.github.io/
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV、cs.GR
Comments NeurIPS 2023 (v3: fixed formula typos in Section 3.5, 43 pages, 34 figures, project page: https://oft.wyliu.com/)
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.GR、cs.MM
Comments 16 pages including appendix
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV、cs.MM
Comments 11 pages, 6 figures
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV、cs.GR
Comments ICCV 2023. Project website: https://www.cs.cmu.edu/~concept-ablation/
专题命中 文生图 :text-to-image(title,abstract);diffusion(title);image generation(abstract);分类 cs.CV、cs.GR
Comments ICCV 2023. Project page at https://orpatashnik.github.io/local-prompt-mixing/