Named-Entity Recognition in the Crime Domain (CrimeNER): Case Study and Dataset
犯罪领域中的命名实体识别(CrimeNER):案例研究与数据集
Miguel Lopez-Duran, Julian Fierrez, Aythami Morales, Daniel DeAlcala, Gonzalo Mancera, Javier Irigoyen, Ruben Tolosana, Oscar Delgado, Francisco Jurado, Alvaro Ortigosa
机构
*
BiometricsAI, Universidad Autónoma de Madrid(生物度AI,马德里自治大学)
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
CommentsThis paper is withdrawn due to significant methodological errors in the experimental design that fundamentally affect the validity of the results. The errors are not correctable within the current framework, and the conclusions can no longer be supported. We apologize for any inconvenience caused to readers
机构
*
The Chinese University of Hong Kong(香港中文大学)
;
Renmin University of China(中国人民大学)
;
Sun Yat-Sen University(中山大学)
;
Chinese Academy of Sciences, Institute of Automation(中国科学院自动化研究所)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
SorryDB: Can AI Provers Complete Real-World Lean Theorems?
SorryDB: AI证明者能完成现实世界的Lean定理吗?
Austin Letson, Leopoldo Sarra, Auguste Poiroux, Oliver Dressler, Paul Lezeau, Dhyan Aranha, Frederick Pu, Aaron Hill, Miguel Corredera Hidalgo, Julian Berman, George Tsoukalas, Lenny Taelman
机构
*
University of California, Berkeley(加州大学伯克利分校)
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.AI、cs.LG
"Do Not Mention This to the User": Detecting and Understanding Malicious Agent Skills in the Wild
“不要向用户提及此事”:检测与理解恶意代理技能
Yi Liu, Zhihao Chen, Yanjun Zhang, Gelei Deng, Yuekang Li, Jianting Ning, Leo Yu Zhang
机构
*
Griffith University(格里菲斯大学)
;
Nanyang Technological University(南洋理工大学)
;
University of New South Wales(新南威尔士大学)
;
Zhejiang Key Laboratory of Digital Fashion and Data Governance, Zhejiang Sci-Tech University(浙江数字时尚与数据治理重点实验室,浙江科技大学)
Illusions of the Gold Standard: A Large-scale Analysis of Human Evaluation Protocols for Long-form Text Generation
黄金标准的幻觉:长文本生成中人类评估协议的大规模分析
Katelyn Xiaoying Mei, Yi-Li Hsu, Minjoon Choi, Zongwan Cao, Chenjun Xu, Bingbing Wen, Su Lin Blodgett, Lucy Lu Wang
机构
*
University of Washington(华盛顿大学)
;
National Tsing Hua University(国立清华大学)
;
Seoul National University(首尔大学)
;
Mila - Québec AI Institute(米拉-魁北克人工智能研究所)
;
Allen Institute for AI(艾伦人工智能研究所)
机构
*
The University of Hong Kong(香港大学)
;
Qwen Team, Alibaba Inc.(阿里巴巴集团Qwen团队)
;
University of California, San Diego(加州大学圣地亚哥分校)
;
Tsinghua University(清华大学)
Comments24 pages, 1 figure. Extended version. A condensed 4-page version appears in the Proceedings of the ACM AI Leadership Summit 2026 (Visionary Papers track)
Patients With Personality: Realistic Patient Simulation through Controlled Diversity and Selective Disclosure
具有个性的患者:通过受控多样性与选择性披露实现逼真的患者模拟
Moritz Schlager, Friederike Jungmann, Samuel Schmidgall, Philipp Raffler, Franziska Hartl, Eva Wende, Paula Roßmüller, Conrad Ketzer, Avinatan Hassidim, Dale R. Webster, Yossi Matias, Yun Liu, Daniel Rueckert, Mike Schaekermann, Paul Hager
机构
*
Technical University of Munich(慕尼黑技术大学)
;
Munich Center for Machine Learning(慕尼黑机器学习中心)
;
TUM University Hospital(慕尼黑技术大学医院)
;
Google DeepMind(谷歌DeepMind)
;
Google Research(谷歌研究)
;
Imperial College London(伦敦帝国学院)
机构
*
Department of Computer and Information Science, University of Pennsylvania(宾夕法尼亚大学计算机与信息科学系)
;
Department of Mathematics, University of Pennsylvania(宾夕法尼亚大学数学系)