AutoMonitor-Bench: Evaluating the Reliability of LLM-Based Misbehavior Monitor
AutoMonitor-Bench:评估基于LLM的误行监控可靠性
机构 * King Abdullah University of Science and Technology(卡塔尔国王 Abdullah 科学与技术大学) ; University of Bristol(布里斯托大学) ; Washington University in St. Louis(圣路易斯华盛顿大学) ; King’s College London(伦敦国王学院) ; Renmin University of China(中国人民大学)
专题命中 代码生成 :code generation(abstract);分类 cs.SE、cs.CL
AI总结 本文提出AutoMonitor-Bench,首个系统评估LLM误行监控可靠性的基准,包含3010个标注样本,通过MR和FAR指标评估12个模型,揭示监控性能的差异及安全与效用的权衡。
Comments ACL 2026 Findings