When Safe Concepts Become Unsafe: Multi-Concept Compositional Vulnerabilities in Text-to-Image Models
TwoHamsters:文本到图像模型中多概念组合不安全性的基准测试
机构 * School of Cyber Science and Engineering, Xi'an Jiaotong University(西安交通大学计算机科学与工程学院) ; CISPA Helmholtz Center for Information Security(信息安全研究中心) ; School of Cyber Science and Engineering, Wuhan University(武汉大学计算机科学与工程学院)
专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV
AI总结 本文提出TwoHamsters基准,通过17500个提示测试文本到图像模型在多概念组合不安全性的表现,揭示现有模型和防御机制在处理危险组合生成时的局限性。