VisualLeakBench: Auditing the Fragility of Large Vision-Language Models against PII Leakage and Social Engineering
VisualLeakBench: 对大型视觉-语言模型在PII泄露和社会工程中的脆弱性进行审计
机构 * Northeastern University(东北大学) ; Carnegie Mellon University(卡内基梅隆大学) ; Boston University(波士顿大学) ; New York University(纽约大学)
专题命中 隐私与版权 :alignment(abstract);safety(abstract);分类 cs.AI
AI总结 本文提出VisualLeakBench,通过合成对抗图像评估LVLMs对OCR注入和PII泄露的鲁棒性,揭示了Claude~4在PII识别上的高风险及防御策略的模板依赖性。