Text-to-image Synthesis via Symmetrical Distillation Networks
专题命中 文生图 :text-to-image(title,abstract);image synthesis(title,abstract);分类 cs.CV
Comments 9 pages, accepted as an oral paper of ACM Multimedia 2018
视觉与机器人
图像生成、文生图、图像编辑、扩散模型和可控生成。
专题命中 文生图 :text-to-image(title,abstract);image synthesis(title,abstract);分类 cs.CV
Comments 9 pages, accepted as an oral paper of ACM Multimedia 2018
专题命中 文生图 :text-to-image(title,abstract);image synthesis(title,abstract);分类 cs.CV
Comments CVPR 2018
专题命中 文生图 :text-to-image(title,abstract);image synthesis(title,abstract);分类 cs.CV
SPOT:基于总变分的选性提示投影用于仅推理的文本到图像生成
机构 * Seoul National University(首尔国立大学)
专题命中 文生图 :text-to-image(title,abstract);image generation(title);diffusion(abstract)
AI总结 本文提出SPOT框架,通过总变分约束生成器参考分布,实现对文本到图像生成中不安全内容的抑制,同时保持良性提示的行为,实验显示其在多个数据集上显著降低不适当评分。
DiffGraph: 一种自动化代理驱动的模型融合框架用于真实场景的文本到图像生成
机构 * Lancaster University(兰卡斯特大学) ; University of Bristol(布里斯托大学) ; Adobe Research(Adobe研究院)
专题命中 文生图 :text-to-image(title,abstract);image generation(title);diffusion(abstract)
AI总结 DiffGraph通过自动化整合在线专家资源,灵活融合不同模型以满足多样化的真实用户需求,提升了文本到图像生成的效能。
Comments CVPR
一个概念远不止一个词:文本到图像扩散模型中的多样化去学习
机构 * VNPT AI(VNPT人工智能研究院) ; VNPT Group(VNPT集团) ; Independent Researcher(独立研究员)
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract)
AI总结 本文提出多样化去学习框架,通过多样化提示集而非单一关键词表示概念,提升去学习的精度与鲁棒性,实验表明其在多个基准和最新基线中表现更优。
CSEval: 一个用于评估文本到图像生成中临床语义的框架
机构 * School of Engineering, University of Edinburgh, United Kingdom(爱丁堡大学工程学院)
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract)
AI总结 CSEval通过语言模型评估文本到图像生成中临床语义的一致性,弥补现有方法在临床相关性评估上的不足。
消除扩散增强交互式文本到图像检索中的幻觉
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract)
AI总结 本文提出DMCL框架,通过引入语义一致性和扩散感知对比目标,有效消除DAI-TIR中的幻觉问题,提升检索性能。
面向文本到图像扩散模型的有效提示窃取攻击
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract)
AI总结 Prometheus通过动态修饰符和上下文匹配算法,实现无需训练的提示窃取攻击,有效提升对多种扩散模型的攻击成功率。
Comments This paper proposes an effective training-free, proxy-in-the-loop, and search-based prompt-stealing scheme against T2I models
SpatialBench-UC: 文本到图像生成中空间提示遵循的不确定性感知评估
机构 * ESIEA
专题命中 文生图 :text-to-image(title,abstract);image generation(title);diffusion(abstract)
AI总结 SpatialBench-UC通过评估文本到图像生成中空间提示遵循的不确定性,提出了一个可重复的基准测试,以提高模型在空间指令下的表现和置信度评估。
Comments 19 pages, includes figures and tables
EnTruth:通过最小和稳健的修改增强文本到图像扩散模型中未经授权数据集使用的可追溯性
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract)
AI总结 EnTruth通过模板记忆技术,提升文本到图像扩散模型中未经授权数据集使用的可追溯性,实现版权保护与生成模型检测的创新应用。
机构 * KAIST(韩国科学技术院) ; NAVER AI Lab(NAVER人工智能实验室) ; SNU AIIS(首尔国立大学人工智能研究所)
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract)
Comments Accepted at NeurIPS 2025
专题命中 文生图 :text-to-image(title,abstract);image synthesis(title,abstract)
Comments Accepted by EMNLP 2025 main
机构 * Minzu University of China(民族大学) ; City University of Macau(澳门城市大学) ; Key Laboratory of Computing Power Network and Information Security, Ministry of Education, Shandong Computer Science Center (National Supercomputer Center in Jinan), Qilu University of Technology (Shandong Academy of Sciences)(计算能力网络与信息安全重点实验室,教育部,山东计算机科学中心(济南国家超级计算机中心),齐鲁工业大学(山东省科学院)) ; The University of Adelaide(阿德莱德大学)
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract)
Comments 16 pages, 13 figures
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract)
专题命中 文生图 :diffusion(title,abstract);image synthesis(title,abstract)
Comments 36 pages, 8 figures, 3 tables, submitted to Elsevier Computerized Medical Imaging and Graphics
专题命中 文生图 :text-to-image(title,abstract);image synthesis(title);diffusion(abstract)
Comments 20 pages. ICCV 2025 (Highlight)
机构 * Department of Computer Science University of Waterloo, Vector Institute(计算机科学系 温哥华大学,向量研究所)
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract)
Comments 17 pages, 6 figures
机构 * TU Darmstadt(图宾根大学) ; Ontocord(Ontocord公司) ; Technical University of Munich (TUM)(慕尼黑技术大学) ; Munich Center for Machine Learning(慕尼黑机器学习中心) ; CERTAIN(CERTAIN公司) ; DFKI(德意志联邦防务研究院) ; Charles University Prague(布拉格查理大学) ; Centre for Cognitive Science, Darmstadt(达姆施塔特认知科学中心) ; Munich Data Science Institute(慕尼黑数据科学研究所)
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract)
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract)
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract)
Comments ICLR 2025; Project Page available at : https://sprain02.github.io/FiFA/
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract)
Comments Submitted to RA-L
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract)
Comments Accepted by ACM CCS 2024
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract)
Comments In "Humane autonomous technology - Re-thinking experience with and in intelligent systems", Palgrave Macmillan, 2024
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract)
Comments Accepted at ICPR 2024
专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract)
Comments Accepted by the 18th European Conference on Computer Vision ECCV 2024
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract)
专题命中 文生图 :text-to-image(title,abstract);image synthesis(title,abstract)
Comments 9 pages, 5 figures. Work in progress
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract)
Comments 5 pages, 4 figures, submitted to 2024 IEEE International Conference on Acoustics, Speech and Signal Processing
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract)
Comments ACM Academic Mindtrek 2023