Detector Guidance for Multi-Object Text-to-Image Generation
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);diffusion(abstract);分类 cs.CV
视觉与机器人
图像生成、文生图、图像编辑、扩散模型和可控生成。
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);diffusion(abstract);分类 cs.CV
专题命中 文生图 :image generation(title,abstract);text-to-image(title);diffusion(abstract);image synthesis(abstract)
Comments 24 pages, 5 figures
专题命中 文生图 :diffusion(title,abstract);image synthesis(title);image generation(abstract);text-to-image(abstract)
专题命中 文生图 :text-to-image(title,abstract);image synthesis(title,abstract);diffusion(abstract);分类 cs.CV
Comments Still on going work
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image generation(abstract);分类 cs.CV
Comments Accepted by CVPR 2023. Project page: https://bluestyle97.github.io/dream3d/
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image generation(abstract);分类 cs.CV
专题命中 文生图 :text-to-image(title,abstract);image synthesis(title,abstract);image generation(abstract);分类 cs.CV
Comments 11 pages, 3 figures
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image synthesis(abstract);分类 cs.CV
Comments The project is available at: https://github.com/Picsart-AI-Research/Text2Video-Zero
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image synthesis(abstract);分类 cs.CV
专题命中 文生图 :text-to-image(title,abstract);image synthesis(title,abstract);diffusion(abstract);分类 cs.CV
Comments 11 pages
专题命中 文生图 :text-to-image(title,abstract);image synthesis(title,abstract);diffusion(abstract);分类 cs.CV
Comments Project page: https://sites.google.com/view/stylegan-t/
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);diffusion(abstract);分类 cs.CV
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image generation(abstract);分类 cs.CV
Comments First Version, 16 pages
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);diffusion(abstract);分类 cs.CV
专题命中 文生图 :text-to-image(title,abstract);image synthesis(title,abstract);image generation(abstract);分类 cs.CV
Comments 14 pages
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);image synthesis(abstract);分类 cs.CV
专题命中 文生图 :text-to-image(title,abstract);image synthesis(title,abstract);image generation(abstract);分类 cs.CV
Comments ICONIP 2021 : The 28th International Conference on Neural Information Processing
Journal ref ICONIP 2021. Lecture Notes in Computer Science, vol 13111, pp 415-426. Springer, Cham
专题命中 文生图 :text-to-image(title,abstract);image synthesis(title,abstract);image generation(abstract);分类 cs.CV
Comments Published at Neural Networks Journal, available at https://www.sciencedirect.com/science/article/pii/S0893608021002823
Journal ref Neural Networks, 2021
专题命中 文生图 :text-to-image(title,abstract);image synthesis(title,abstract);image generation(abstract);分类 cs.CV
Comments Accepted at ACM MM 2021
文本到图像模型中的明喻理解:一个评估框架
机构 * The University of Tokyo(东京大学) ; Nara Institute of Science and Technology(奈良科学技术研究所) ; Chungnam National University(忠南国立大学) ; Institute of Science Tokyo(东京科学大学)
专题命中 文生图 :diffusion(summary_cn,abstract);text-to-image(title,abstract);分类 cs.CV、cs.MM
AI总结 针对文本到图像模型常混淆明喻喻体与本体的问题,提出含受控数据集、YOLO指标及Diffusion Lens分析的评估框架,实验发现模型存在字面化失败模式并讨论了缓解策略。
Comments Accepted as a full paper at ACM Multimedia 2026
LegoDiffusion:微服务文本到图像扩散工作流
机构 * Hong Kong University of Science and Technology(香港科技大学) ; Alibaba Group(阿里巴巴集团)
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image generation(abstract)
AI总结 LegoDiffusion通过将扩散工作流分解为松耦合的模型执行节点,实现高效管理与调度,提升请求率和突发流量容忍度。
文本到图像扩散模型的持续反学习:一种正则化视角
机构 * The Ohio State University(俄亥俄州立大学) ; Michigan State University(密歇根州立大学) ; Texas A&M University(德克萨斯大学安德森分校) ; Boston University(波士顿大学)
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image generation(abstract)
AI总结 本文提出通过正则化方法解决文本到图像扩散模型中持续反学习的问题,通过减轻参数漂移和增强语义意识,提升反学习性能并推动生成式AI的安全发展。
Comments Accepted to ICLR 2026
SafeGen: 在文本到图像生成中嵌入伦理保障
机构 * Posts and Telecommunications Institute of Technology(电信技术研究所)
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);diffusion(abstract)
AI总结 SafeGen通过整合BGE-M3和Hyper-SD,实现文本到图像生成中的伦理保障,有效过滤有害提示并生成高质量图像。
机构 * KAIST(韩国科学技术院) ; NAVER Cloud(NAVER云)
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image editing(abstract)
Comments Accepted at NeurIPS 2025
机构 * KAIST(韩国科学技术院)
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image generation(abstract)
Comments 29 pages, Under review
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image generation(abstract)
机构 * WMG, The University of Warwick(沃里克大学工学院) ; Dept. of Computer Science, University of Liverpool(利物浦大学计算机科学系) ; Dept. of Computer Science & Engineering, Chalmers University of Technology(楚科奇斯技术大学计算机科学与工程系) ; James-Watt Engineering School, University of Glasgow(格拉斯哥大学詹姆斯-沃特工程学院)
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image generation(abstract)
Comments under review
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image generation(abstract)
Comments 10 pages, 10 figures, Mechanistic Interpretability for Vision at CVPR 2025
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image generation(abstract)
Comments 16 pages, 9 figures
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);diffusion(abstract)