MultiBanana: A Challenging Benchmark for Multi-Reference Text-to-Image Generation
MultiBanana:多参考文本到图像生成的挑战性基准
机构 * The University of Tokyo(东京大学) ; Google DeepMind(谷歌DeepMind)
专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);分类 cs.CV
AI总结 本文提出MultiBanana基准,用于评估多参考文本到图像生成模型的能力,涵盖参考数量变化、领域不匹配、尺度不匹配、罕见概念和多语言参考等挑战,分析不同模型的性能与不足。
Comments Accepted to CVPR2026