机构
*
Southern University of Science and Technology(南方科技大学)
;
University of Science and Technology of China(中国科学技术大学)
;
University of Birmingham(伯明翰大学)
;
Zhejiang University(浙江大学)
;
East China Normal University(华东师范大学)
;
Alibaba Group(阿里巴巴集团)
Systematic Evaluation of the Quality of Synthetic Clinical Notes Rephrased by LLMs at Million-Note Scale
在百万笔记规模上系统评估LLM重新表述的合成临床笔记质量
Jinghui Liu, Sarvesh Soni, Anthony Nguyen
机构
*
Australian e-Health Research Centre, CSIRO, Australia(澳大利亚电子健康研究中心,CSIRO,澳大利亚)
;
National Library of Medicine, National Institutes of Health, USA(国家医学图书馆,国立卫生研究院,美国)
专题命中
评测与基准
:LLM(title_cn,summary_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
Auditing Multimodal LLM Raters: Central Tendency Bias in Clinical Ordinal Scoring
对多模态大语言模型评分者的审计:临床顺序评分中的中间倾向偏差
Jiaqing Zhang, Sandeep Elluri, Bhanu Cherukuvada, Yonah Joffe, Jessica Sena, Miguel Contreras, Scott Siegel, Subhash Nerella, Catherine Price, Parisa Rashidi
机构
*
Department of Electrical & Computer Engineering(电气与计算机工程系)
;
Department of Computer and Information Science and Engineering(计算机与信息科学与工程系)
;
Department of Clinical and Health Psychology(临床与健康心理学系)
;
Department of Biomedical Engineering(生物医学工程系)
专题命中
评测与基准
:LLM(title,summary_cn);large language model(abstract);language model(abstract)
机构
*
Northeastern University(东北大学)
;
University of Southern California(南加州大学)
;
Stony Brook University(石溪大学)
;
Independent Researcher(独立研究者)
;
Ohio State University(俄亥俄州立大学)
;
University of Notre Dame(Notre Dame 大学)
;
Columbia University(哥伦比亚大学)
专题命中
评测与基准
:LLM(title_cn,summary_cn);large language model(abstract);language model(abstract);分类 cs.CL
HalluScore: Large Language Model Hallucination Question Answering Benchmark
HalluScore: 大语言模型幻觉问答基准
Aisha Alansari, Hamzah Luqman
机构
*
Department of Information and Computer Science, King Fahd University of Petroleum and Minerals(国王法赫德石油与矿物大学信息与计算机科学系)
;
SDAIA-KFUPM Joint Research Center for Artificial Intelligence(SDAIA-KFUPM人工智能联合研究中心)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);LLM(abstract_cn);分类 cs.CL
机构
*
School of Economics and Management, East China Normal University(东华大学经济管理学院)
;
School of Information Management, Wuhan University(武汉大学信息管理学院)
;
China Academic Degrees & Graduate Education Development Center(中国学位与研究生教育发展中心)
专题命中
评测与基准
:LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL
机构
*
Georgia Institute of Technology(佐治亚理工学院)
;
Columbia University(哥伦比亚大学)
;
California State University(加州州立大学)
;
University of Montreal(蒙特利尔大学)
;
Carnegie Mellon University(卡内基梅隆大学)
;
Rensselaer Polytechnic Institute(莱斯利理工学院)
;
The University of Manchester(曼彻斯特大学)
;
Harvard University(哈佛大学)
专题命中
评测与基准
:LLM(title_cn,summary_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
机构
*
College of Computer Science and Software Engineering, Shenzhen University(深圳大学计算机科学与软件工程学院)
;
Microsoft Research Asia(微软亚洲研究院)
;
Microsoft Gaming(微软游戏)
;
Provincial Key Laboratory of Intelligent Communication and Digital Society Governance, Shenzhen University(深圳大学省级智能通信与数字社会治理重点实验室)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);分类 cs.AI
From Pixels to Places: A Systematic Benchmark for Evaluating Image Geolocalization Ability in Large Language Models
从像素到地点:一个系统性基准,用于评估大语言模型中的图像地理定位能力
Lingyao Li, Runlong Yu, Qikai Hu, Bowei Li, Min Deng, Yang Zhou, Xiaowei Jia
机构
*
University of South Florida Tampa USA
;
University of Alabama Tuscaloosa USA
;
University of Michigan Ann Arbor USA
;
Texas Tech University Lubbock USA
;
Texas A \& M University College Station USA
;
University of Pittsburgh Pittsburgh USA
;
University of South Florida
;
University of Alabama
;
University of Michigan
;
Texas Tech University
;
Texas A \& M University
;
University of Pittsburgh
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract)
机构
*
School of IC, Peking University(北京大学集成电路学院)
;
School of EECS, Peking University(北京大学电子信息技术学院)
;
School of Microelectronics, Xidian University(西安电子科技大学微电子学院)
;
Institute of EDA, Peking University(北京大学EDA研究院)
;
Beijing Advanced Innovation Center for IC(北京集成电路先进创新中心)
Trustworthiness in Retrieval-Augmented Generation Systems: A Survey
检索增强生成系统中的可信度:综述
Yujia Zhou, Wenbo Zhang, Jingying Shao, Yan Liu, Xiaoxi Li, Jiajie Jin, Hongjin Qian, Zheng Liu, Chaozhuo Li, Jason Chen Zhang, Zhicheng Dou, Philip S. Yu, Jiaxin Mao
机构
*
Tsinghua University(清华大学)
;
Renmin University of China(中国人民大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Beijing Academy of Artificial Intelligence(北京人工智能研究院)
;
Hong Kong Polytechnic University(香港理工大学)
;
Microsoft Research Asia(微软亚洲研究院)
;
University of Illinois(伊利诺伊大学)
专题命中
评测与基准
:LLM(summary_cn,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
Evidence-Grounded Frontier Mapping and Agentic Hypothesis Generation in Nanomedicine
基于证据的前沿映射与代理假设生成在纳米医学中
Christiaan G. A. Viviers, Koen de Bruin, Mirre M. Trines, Ayla M. Hokke, Roy van der Meel, Avi Schroeder, Twan Lammers, Willem J. M. Mulder, Fons van der Sommen
机构
*
ARIA Lab, Signal Processing Systems, Department of Electrical Engineering, Eindhoven University of Technology(ARIA实验室,信号处理系统,电气工程系,埃因霍温理工大学)
;
Laboratory of Chemical Biology, Department of Biomedical Engineering, Eindhoven University of Technology(化学生物学实验室,生物医学工程系,埃因霍温理工大学)
;
The Louis Family Laboratory for Targeted Drug Delivery and Personalized Medicine Technologies, Department of Chemical Engineering, Technion - Israel Institute of Technology(定向药物输送与个性化医学技术实验室,化学工程系,技术离子-以色列理工学院)
;
Department of Nanomedicine and Theranostics, Institute for Experimental Molecular Imaging (ExMI), RWTH Aachen University Hospital(纳米医学与诊疗学系,实验分子成像研究所(ExMI),亚琛工业大学医院)
;
Department of Internal Medicine and Radboud Center for Infectious Diseases (RCI), Radboud University Medical Center(内科学系和Radboud感染疾病中心(RCI),Radboud大学医学中心)
专题命中
评测与基准
:LLM(summary_cn,abstract);large language model(abstract);language model(abstract);分类 cs.AI