Simile Understanding in Text-to-Image Models: An Evaluation Framework
文本到图像模型中的明喻理解:一个评估框架
机构 * The University of Tokyo(东京大学) ; Nara Institute of Science and Technology(奈良科学技术研究所) ; Chungnam National University(忠南国立大学) ; Institute of Science Tokyo(东京科学大学)
专题命中 文生图 :diffusion(summary_cn,abstract);text-to-image(title,abstract);分类 cs.CV、cs.MM
AI总结 针对文本到图像模型常混淆明喻喻体与本体的问题,提出含受控数据集、YOLO指标及Diffusion Lens分析的评估框架,实验发现模型存在字面化失败模式并讨论了缓解策略。
Comments Accepted as a full paper at ACM Multimedia 2026