ANCHOR: LLM-driven Subject Conditioning for Text-to-Image Synthesis
ANCHOR:基于大语言模型的文本到图像合成中的主体条件化
机构 * Optum AI ; The Pennsylvania State University(宾夕法尼亚州立大学)
专题命中 文生图 :text-to-image(title,abstract);image synthesis(title,abstract);分类 cs.CV、cs.MM
AI总结 ANCHOR通过大规模抽象式标题数据集研究文本到图像合成中多主体理解与上下文推理的缺陷,提出基于大语言模型的主体感知微调方法,提升图像-标题一致性与人类偏好对齐。
Comments Accepted to The 64th Annual Meeting of the Association for Computational Linguistics (ACL) 2026