Trojan Horses in Recruiting: A Red-Teaming Case Study on Indirect Prompt Injection in Standard vs. Reasoning Models
招聘中的木马:针对标准与推理模型间接提示注入的红队案例研究
机构 * University of Mannheim(曼海姆大学)
专题命中 逻辑推理 :reasoning(title,abstract);chain-of-thought(abstract);分类 cs.AI
AI总结 本研究通过红队测试揭示了标准与推理模型在间接提示注入中的安全差异,发现推理模型在复杂指令下易出现元认知泄漏,而标准模型在简单攻击中表现较弱。
Comments 43 pages, 3 synthetic CV PDF's, 6 chat history PDF's and system prompts. This work was developed as part of the Responsible AI course within the Mannheim Master in Data Science (MMDS) program at the University of Mannheim