Reviewing Scientific Papers for Critical Problems With Reasoning LLMs: Baseline Approaches and Automatic Evaluation
利用推理LLM审查科学论文中的关键问题:基线方法与自动评估
机构 * University of Washington(华盛顿大学)
专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL
AI总结 本文提出利用推理LLM作为论文质量检查器,介绍基线方法和自动评估框架,通过arXiv论文验证方法,并评估其在识别科学论文关键错误中的性能。
Comments Accepted and presented at NeurIPS 2025 AI for Science Workshop