Consistent but Miscalibrated: Evaluating LLM Limitations for Risk Communication in Natural Language
一致但校准错误:评估大语言模型在自然语言风险沟通中的局限性
Diego Cerda-Mardini, Sarath Chandar, Sreenath Madathil
机构
*
Faculty of Dental Medicine and Oral Health Sciences, McGill University(麦吉尔大学牙医学与口腔健康科学学院)
;
Chandar Research Lab, Polytechnique Montréal(蒙特利尔理工大学钱达尔研究实验室)
;
Mila – Québec Artificial Intelligence Institute(魁北克人工智能研究所米拉)
;
Département de Génie Informatique et Génie Logiciel (GIGL), Polytechnique Montréal(蒙特利尔理工大学计算机科学与软件工程系)
Knowledge-Based Design Requirements for Generative Social Robots in Higher Education
面向高等教育的生成社交机器人知识基础设计需求
Stephan Vonschallen, Dominique Oberle, Theresa Schmiedel, Friederike Eyssel
机构
*
Zurich University of Applied Sciences(应用科学大学苏黎世)
;
University of Applied Sciences and Arts Northwestern Switzerland(西北瑞士应用科学与艺术大学)
;
Bielefeld University(比勒菲尔德大学)
专题命中
领域大模型
:large language model(abstract);language model(abstract);分类 cs.AI
Shortcut Learning in Legal Judgment Prediction: Empirical Evidence from the UK Employment Tribunal
法律判决预测中的捷径学习:来自英国就业法庭的实证证据
Joe Watson, Joana Ribeiro de Faria, Marcus Tomalin, Måns Magnusson, Huiyuan Xie, Hao Tian Yeung, Christine Carter, Jonathan Rutherford, Felix Steffek
机构
*
Faculty of Law, University of Cambridge(剑桥大学法学院)
;
The Psychometrics Centre, Cambridge Judge Business School, University of Cambridge(剑桥大学心理学测量中心,剑桥Judge商学院)
;
Faculty of English, University of Cambridge(剑桥大学英语学院)
;
Department of Statistics, Uppsala University(乌普萨拉大学统计系)
;
Department of Computer Science and Technology, Tsinghua University(清华大学计算机科学与技术系)
;
Department of Engineering, University of Cambridge(剑桥大学工程系)
机构
*
Institute of AI for Health & Helmholtz AI, Computational Health Center, Helmholtz Munich – German Research Center for Environmental Health(亥姆霍兹慕尼黑中心 - 德国环境健康研究中心,计算健康中心,健康人工智能研究所与亥姆霍兹人工智能)
;
Department of Medicine III, Ludwig-Maximilian-University Hospital(路德维希-马克西米利安大学医院第三内科)
;
Department of Physics, Ludwig-Maximilian-University(路德维希-马克西米利安大学物理系)
;
German Cancer Consortium (DKTK), partner site Munich(德国癌症联盟(DKTK)慕尼黑合作站点)
;
Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心(MCML))
LUMOS: Latent Universal Medical Priors for Segmentation
GuiDINO:重新思考医学图像分割中的视觉基础模型
Zhuonan Liang, Wei Guo, Jie Gan, Yaxuan Song, Runnan Chen, Hang Chang, Weidong Cai
机构
*
The University of Sydney, Sydney. NSW 2006, Australia Biological System \& Engineering Division, Lawrence Berkeley National Laboratory, Berkeley. CA 94720, USA Berkeley Biomedical Data Science Center, Lawrence Berkeley National Laboratory, Berkeley. CA 94720, USA
CommentsAccepted at the AAAI 2026 Summer Symposium Series. This paper was presented at the ICLR Re-Align workshop under a different name, "Accelerating Adversarial Suffix Optimization via Continuous Relaxation and Activation-Guided Objectives"
机构
*
School of Airspace and Engineering, Shandong University(山东大学航空航天与工程学院)
;
School of Computer Science and Technology, Shandong University of Finance and Economics(山东财经大学计算机科学与技术学院)
;
School of Psychology and Neuroscience, University of Glasgow(格拉斯哥大学心理学与神经科学学院)
;
Faculty of Information Science and Engineering, Ocean University of China(中国海洋大学信息科学与工程学院)
LLM for EDA in Front-End Design: Challenges and Opportunities
用于前端设计中电子设计自动化的大语言模型:挑战与机遇
Kangwei Xu, Bing Li, Ulf Schlichtmann
机构
*
Chair of Electronic Design Automation, Technical University of Munich(电子设计自动化教授团,慕尼黑技术大学)
;
Resource-Efficient AI Group, Technical University of Ilmenau(高效人工智能小组,伊门豪技术大学)
专题命中
其他LLM
:LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.LG
Comments25 pages, 7 figures. ProofCouncil appears as System A (IMProofBench ProofCouncil) in the official FirstProof second-batch report (arXiv:2606.18119). Code and agent-building library: https://github.com/eth-sri/proof-council