Verbalizable Representations Form a Global Workspace in Language Models
语言模型中的可言语化表征形成全局工作空间
Wes Gurnee, Nicholas Sofroniew, Adam Pearce, Mateusz Piotrowski, Isaac Kauvar, Runjin Chen, Anna Soligo, Paul Bogdan, Euan Ong, Rowan Wang, Ben Thompson, David Abrahams, Subhash Kantamneni, Emmanuel Ameisen, Joshua Batson, Jack Lindsey
机构
*
Universidad de Buenos Aires, Facultad de Ciencias Exactas y Naturales, Departamento de Computación(布宜诺斯艾利斯大学,精确与自然科学学院,计算机系)
;
AI Safety Argentina (AISAR)(阿根廷人工智能安全组织 (AISAR))
;
Department of Computer Science, University of Oxford(牛津大学计算机科学系)
;
CONICET-Universidad de Buenos Aires, Instituto de Ciencias de la Computación (ICC)(阿根廷国家科学与技术研究理事会-布宜诺斯艾利斯大学,计算机科学研究所 (ICC))
From Knowing to Acting: Benchmarking Self-Awareness Capability of LLM Agents
从知道到行动:基准测试LLM代理的自我意识能力
Yifan Li, Shengbin Yue, Boyu Feng, Jinhu Qi, Bo Ke, Zixing Song, Hongru Wang, Zhongyu Wei, Irwin King
机构
*
The Chinese University of Hong Kong(香港中文大学)
;
Fudan University(复旦大学)
;
University of Edinburgh(爱丁堡大学)
;
Tencent(腾讯)
;
University of Bristol(布里斯托大学)
On the Adversarial Robustness of Multimodal LLM Judges
多模态大语言模型评判器的对抗鲁棒性
Zihan Wang, Guansong Pang, Zelin Liu, Wenjun Miao, Jin Zheng, Xiao Bai
机构
*
School of Computer Science and Engineering, Beihang University(北京航空航天大学计算机科学与工程学院)
;
State Key Laboratory of Virtual Reality Technology and System, Beihang University(北京航空航天大学虚拟现实技术与系统国家重点实验室)
;
State Key Laboratory of Software Development Environment, Jiangxi Research Institute, Beihang University(北京航空航天大学江西研究院软件开发环境国家重点实验室)
;
School of Computing and Information Systems, Singapore Management University(新加坡管理大学计算机与信息系统学院)
OpenMedReason: Scientific Reasoning Supervision for Medical Vision-Language Models
OpenMedReason: 医学视觉语言模型的科学推理监督
Negin Baghbanzadeh, Pritam Sarkar, Michael Colacci, Abeer Badawi, Adibvafa Fallahpour, Arash Afkanpour, Leonid Sigal, Ali Etemad, Elham Dolatabadi
机构
*
York University(约克大学)
;
Vector Institute(向量研究所)
;
University of British Columbia(不列颠哥伦比亚大学)
;
University of Toronto(多伦多大学)
;
Unity Health Toronto / St. Michael’s Hospital(多伦多联合健康/圣迈克尔医院)
;
University Health Network(大学健康网络)
;
Arc Institute(弧研究所)
;
Queen's University(女王大学)
PreAct-Bench: Benchmarking Predictive Monitoring in LLMs
PreAct-Bench:大语言模型中的预测性监控基准
Hainiu Xu, Italo Luis da Silva, Jiangnan Ye, Yuhao Wang, Wei Liu, Linyi Yang, Jonathan Richard Schwarz, Nicola Paoletti, Yulan He, Hanqi Yan
机构
*
King’s College London(伦敦国王学院)
;
National University of Singapore(新加坡国立大学)
;
Southern University of Science and Technology(南方科技大学)
;
Thomson Reuters Foundational Research(汤姆森路透基础研究)
;
Imperial College London(伦敦帝国学院)
;
The Alan Turing Institute(艾伦·图灵研究所)
CommentsICAHS, \c{opyright} 2025 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work in other works
Subtitle-Aligned Fine-Tuning of Whisper for Swiss German ASR: Benchmark Contamination, Convention Mismatch, and an Honest Baseline at 25.6% WER (13.8% cWER)
CommentsProceedings of the 6th Workshop on Trustworthy NLP (TrustNLP 2026), ACL 2026, San Diego, California, USA. Available at https://openreview.net/forum?id=WJCalficPT
BELLS-O: Evaluating the Operational Trade-offs of LLM Supervision Systems
BELLS-O:评估LLM监督系统的运营权衡
Leonhard Waibl, Felix Michalak, Hadrien Mariaccia
机构
*
University of Graz, Graz, Austria(格拉茨大学)
;
Supervised Program for Alignment Research (SPAR)(对齐研究监督计划 (SPAR))
;
Centre pour la Sécurité de l'IA (CeSIA), Paris, France(人工智能安全研究中心 (CeSIA),巴黎,法国)