MEDIC: Comprehensive Evaluation of Leading Indicators for LLM Safety and Utility in Clinical Applications
MEDIC:对LLM在临床应用中的安全性和实用性领先指标的综合评估
Praveenkumar Kanithi, Clément Christophe, Marco AF Pimentel, Tathagata Raha, Prateek Munjal, Nada Saadi, Hamza A Javed, Svetlana Maslenkova, Nasir Hayat, Ronnie Rajan, Shadab Khan
Unsupervised Deep Learning for Inverse Problems in Computed Tomography
计算断层成像逆问题的无监督学习
Laura Hellwege, Johann Christopher Engster, Moritz Schaar, Thorsten M. Buzug, Maik Stille
机构
*
Institute of Medical Engineering, University of Lübeck(吕贝克大学医学工程研究所)
;
Fraunhofer Research Institution for Individualized Medical Technology and Engineering (IMTE)(弗劳恩霍夫个性化医疗技术与工程研究所)
Pipette: An Embodied Simulation Platform, Benchmark, and Data-Efficient Augmentation Framework for Wet-Lab Robotics
面向湿实验室机器人的具身仿真平台、基准测试及数据高效增强框架
Zhe Liu, Huanbo Jin, Zhaohui Du, Zhe Wang, Dongzhan Zhou, Minting Pan, He Xu, Peijia Li, Jiaming Gu, Quan Lu, Qi Wang, Bin Ji, Ting Xiao
机构
*
Key Laboratory of Smart Manufacturing in Energy Chemical Process Ministry of Education(能源化工过程智能制造国家重点实验室)
;
Department of Computer Science and Engineering(计算机科学与工程系)
;
Department of Laboratory Medicine(实验室医学系)
;
Shanghai Jiao Tong University School of Medicine(上海交通大学医学院)
Health System Scale Semantic Search Across Unstructured Clinical Notes
跨无结构临床笔记的健康系统规模语义搜索
Faith Wavinya Mutinda, Spandana Makeneni, Anna Lin, Shivaji Dutta, Irit R. Rasooly, Patrick Dibussolo, Shivani Kamath Belman, Hessam Shahriari, Kevin Murphy, Alex B. Ruan, Barbara H. Chaiyachati, Sanjay Chainani, Robert W. Grundmeier, Scott M. Haag, Jeffrey M. Miller, Heather M. Griffis, Ian M. Campbell
机构
*
Department of Biomedical and Health Informatics, Children’s Hospital of Philadelphia(儿童医院哲学学院生物医学与健康信息学系)
;
Google Cloud(谷歌云)
;
Department of Pediatrics, University of Pennsylvania(宾夕法尼亚大学儿科系)
;
Division of Neonatology, Children’s Hospital of Philadelphia(儿童医院哲学学院新生儿科)
;
Division of Human Genetics, Children’s Hospital of Philadelphia(儿童医院哲学学院人类遗传学部)
AgenticRAGTracer: A Hop-Aware Benchmark for Diagnosing Multi-Step Retrieval Reasoning in Agentic RAG
AgenticRAGTracer:用于诊断Agentic RAG中多步检索推理的跳数感知基准
Qijie You, Wenkai Yu, Wentao Zhang
机构
*
University of Science and Technology Beijing(北京科技大学)
;
Peking University(北京大学)
;
Zhongguancun Academy(中关村学院)
;
Beijing Key Laboratory of Data Intelligence and Security (Peking University)(北京市数据智能与安全重点实验室(北京大学))
MedLayBench-V: A Large-Scale Benchmark for Expert-Lay Semantic Alignment in Medical Vision Language Models
MedLayBench-V:面向医学视觉语言模型中专家与普通人语义对齐的大规模基准
Han Jang, Junhyeok Lee, Heeseong Eum, Kyu Sung Choi
机构
*
Seoul National University(首尔国立大学)
;
Seoul National University College of Medicine(首尔国立大学医学院)
;
Department of Radiology, Seoul National University Hospital(首尔国立大学医院放射科)
;
Healthcare AI Research Institute, Seoul National University Hospital(首尔国立大学医院健康人工智能研究所)
;
The Advanced Imaging and Computational Neuroimaging (AICON) Laboratory(先进影像与计算神经影像实验室)
机构
*
Department of Computer Science, University of the Cumberlands(大学的计算机科学系)
;
Department of Computer Science, DePaul University(德保罗大学计算机科学系)
;
Youngstown State University(亚当斯州立大学)
Large Language Models for Agentic NetOps and AIOps: Architectures, Evaluation, and Safety
用于代理网络运维和AI运维的大型语言模型:架构、评估与安全
Muhammad Bilal, Jon Crowcroft, Ruizhi Wang, Xiaolong Xu, Schahram Dustdar
机构
*
School of Computing and Communications(计算与通信学院)
;
University of Cambridge(剑桥大学)
;
School of Software(软件学院)
;
Nanjing University of Information Science and Technology(南京信息科技大學)
;
TU Wien(维也纳技术大学)
;
ICREA
Noisy-QSMOTE: Robustness Analysis of Quantum SMOTE under Quantum-Inspired Noise for Condition Monitoring and Fault Classification in Industrial and Energy Systems