RE-MCDF: Closed-Loop Multi-Expert LLM Reasoning for Knowledge-Grounded Clinical Diagnosis
RE-MCDF:闭环多专家LLM推理用于知识驱动的临床诊断
Shaowei Shen, Xiaohong Yang, Jie Yang, Lianfen Huang, Yongcai Zhang, Yang Zou, Seyyedali Hosseinalipour
机构
*
School of Informatics, Xiamen University, China(厦门大学信息学系, 中国)
;
National Institute for Data Science in Health and Medicine, Xiamen University, China(健康医学数据科学国家研究院, 厦门大学, 中国)
;
Key Laboratory of Intelligent Manufacturing Equipment and Industrial Internet Technology, Fujian Provincial Universities, the School of Information Science and Technology, Xiamen University Tan Kah Kee College, and also with the Department of Informatics and Communication Engineering, Xiamen University, China(福建省智能制造装备与工业互联网技术重点实验室, 福建省高校, 厦门大学信息科学系, 厦门大学坦克 Kee 学院, 以及厦门大学信息与通信工程系, 中国)
;
School of Medicine, Xiamen University, China(厦门大学医学院, 中国)
;
School of Electronic and Information Engineering, Tongji University, China(同济大学电子与信息工程学院, 中国)
;
department of Electrical Engineering, University at Buffalo-SUNY, Buffalo, NY, USA(University at Buffalo-SUNY 电气工程系, Buffalo, NY, 美国)
A Survey on the Safety and Security Threats of Computer-Using Agents: JARVIS or Ultron?
对计算机使用代理的安全性和安全威胁的综述:贾维斯或乌tron?
Ada Chen, Yongjiang Wu, Junyuan Zhang, Jingyu Xiao, Shu Yang, Jen-tse Huang, Kun Wang, Wenxuan Wang, Shuai Wang
机构
*
Carnegie Mellon University(卡内基梅隆大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
KAUST(卡塔尔科技大学)
;
Johns Hopkins University(约翰霍普金斯大学)
;
Nanyang Technological University(南洋理工大学)
;
Renmin University of China(中国人民大学)
;
The Hong Kong University of Science and Technology(香港科学大学)
机构
*
Gaoling School of Artificial Intelligence, Renmin University of China, Beijing, China(中国人民大学北京校区人工智能学院)
;
DAMO Academy, Alibaba Group, Hangzhou, China(阿里巴巴集团达摩院,杭州,中国)
;
Hupan Lab, Hangzhou, China(虎扑实验室,杭州,中国)
;
Institution of Physics, University of the Chinese Academy of Sciences, Beijing, China(中国科学院物理研究所,北京,中国)
;
Department of Computer Science and Technology, Tsinghua University, Beijing, China(清华大学计算机科学与技术系,北京,中国)
Value-Guided Iterative Refinement and the DIQ-H Benchmark for Evaluating VLM Robustness
基于价值引导的迭代精炼与DIQ-H基准:评估VLM鲁棒性的新方法
Hanwen Wan, Zexin Lin, Yixuan Deng, Xiaoqiang Ji
机构
*
The School of Science and Engineering, The Chinese University of Hong Kong, Shenzhen, China(香港中文大学(深圳)科学与工程学院)
;
The School of Artificial Intelligence, The Chinese University of Hong Kong, Shenzhen, China(香港中文大学(深圳)人工智能学院)
;
The Shenzhen Institute of Artificial Intelligence and Robotics for Society, Shenzhen, China(深圳人工智能与机器人社会研究院)
AI Observability for Large Language Model Systems: A Multi-Layer Analysis of Monitoring Approaches from Confidence Calibration to Infrastructure Tracing
Journal refProceedings of the 49th International ACM SIGIR Conference on Research and Development in Information Retrieval (SIGIR '26), July 20--24, 2026, Melbourne, VIC, Australia
Through a Compressed Lens: Investigating The Impact of Quantization on Factual Knowledge Recall
通过压缩镜头:探讨量化对事实知识回忆的影响
Qianli Wang, Mingyang Wang, Nils Feldhus, Simon Ostermann, Yuan Cao, Hinrich Schütze, Sebastian Möller, Vera Schmitt
机构
*
Quality and Usability Lab, Technische Universität Berlin(柏林技术大学质量与可用性实验室)
;
German Research Center for Artificial Intelligence (DFKI)(德国人工智能研究中心)
;
Saarland Informatics Campus(萨尔州信息学院)
;
LMU Munich(慕尼黑大学)
;
Bosch Center for Artificial Intelligence (BCAI)(博世人工智能中心)
;
Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心)
;
Centre for European Research in Trusted AI (CERTAIN)(可信人工智能欧洲研究中心)
;
BIFOLD – Berlin Institute for the Foundations of Learning and Data(柏林学习与数据基础研究院)
;
Technical University of Munich(慕尼黑技术大学)
Identifying the Achilles' Heel: An Iterative Method for Dynamically Uncovering Factual Errors in Large Language Models
找出阿基里斯之踵:一种用于动态揭示大语言模型事实错误的迭代方法
Wenxuan Wang, Yuk-Kit Chan, Zixuan Ling, Juluan Shi, Youliang Yuan, Jen-tse Huang, Yifei Zhang, Wenxiang Jiao, Zhaopeng Tu, Michael R. Lyu
机构
*
Renmin University of China(中国人民大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学深圳校区)
;
Johns Hopkins University(约翰霍普金斯大学)
;
Nanyang Technological University(南洋理工大学)
;
Xiaohongshu Inc.(小红书公司)
;
Tencent Inc.(腾讯公司)
Can LLMs Help Allocate Public Health Resources? A Case Study on Childhood Lead Testing
大语言模型能否帮助分配公共卫生资源?一项关于儿童铅检测的案例研究
Mohamed Afane, Ying Wang, Juntao Chen
机构
*
Department of Computer and Information Sciences, Fordham University(福特汉姆大学计算机与信息科学系)
;
Department of Systems Engineering, Stevens Institute of Technology(史蒂文斯理工学院系统工程系)