Using LLM-as-a-Judge/Jury to Advance Scalable, Clinically-Validated Safety Evaluations of Model Responses to Users Demonstrating Psychosis
利用LLM作为法官/陪审团推进模型响应用户表现精神病的可扩展、临床验证的安全性评估
机构 * Apart Research ; Odyssean Institute ; London School of Economics and Political Science(伦敦政治经济学院) ; Stanford Institute for Human-Centered AI(斯坦福大学以人为本人工智能研究所)
专题命中 安全评测 :safety(title,abstract);分类 cs.CL、cs.AI
AI总结 本文提出利用LLM作为法官或陪审团评估模型响应用户精神病表现的方法,开发了七项临床验证的安全标准,并展示了LLM作为法官在临床验证中的有效性。
Comments published at IASEAI 2026, preliminary work presented at GenAI4Health workshop at NeurIPS 2025