CommentsMain changes: - Slightly altered title & author ordering - New section detailing survey methodology - Expanded literature coverage and improved discussion of all references for clarity, precision & conciseness - Removed the "appealing to authority" subsection & integrated its content elsewhere - Overhauled the experimental design section - Significantly expanded success metrics discussion
Agentic Large Language Models for Training-Free Neuro-Radiological Image Analysis
基于代理的大型语言模型用于无训练神经放射学图像分析
Ayhan Can Erdur, Daniel Scholz, Jiazhen Pan, Benedikt Wiestler, Daniel Rueckert, Jan C. Peeken
机构
*
Department of Radiation Oncology, TUM University Hospital, Munich, Germany(放射肿瘤科,慕尼黑技术大学医院,德国)
;
Chair for AI in Healthcare and Medicine, Technical University of Munich (TUM)(人工智能在医疗和健康领域的主任,慕尼黑技术大学)
;
Chair for AI for Image-Guided Diagnosis and Therapy, Technical University of Munich (TUM)(人工智能在影像引导诊断和治疗领域的主任,慕尼黑技术大学)
;
Munich Center for Machine Learning (MCML), Munich, Germany(慕尼黑机器学习中心(MCML),德国慕尼黑)
;
Department of Computing, Imperial College London, London, UK(计算系,伦敦帝国学院,英国伦敦)
;
Deutsches Konsortium für Translationale Krebsforschung (DKTK), Partner Site Munich, Munich, Germany(德国转化癌症研究联盟(DKTK),慕尼黑分部,德国慕尼黑)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);LLM(abstract_cn);分类 cs.AI
Large language models for post-publication research evaluation: Evidence from expert recommendations and citation indicators
基于大语言模型的发表后研究评估:专家推荐与引用指标的证据
Mengjia Wu, Yi Zhang, Robin Haunschild, Lutz Bornmann
机构
*
Australian Artificial Intelligence Institute, Faculty of Engineering and Information Technology, University of Technology Sydney(澳大利亚人工智能研究所,工程与信息科技学院,新南威尔士大学)
;
Max Planck Institute for Solid State Research(马克斯·普朗克固体物理研究所)
;
Science Policy and Strategy Department, Administrative Headquarters of the Max Planck Society(马克斯·普朗克学会科学政策与战略部门,行政总部)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);prompting(abstract);分类 cs.AI
CommentsA benchmark for evaluating multimodal both voice and text LLM agents in dualcontrol settings. We introduce persona adaptive prompting and 12 new metrics to assess robustness safety efficiency and recovery in customer support scenarios
FDM-Bench: A Comprehensive Benchmark for Evaluating Large Language Models in Additive Manufacturing Tasks
FDM-Bench:用于评估大型语言模型在增材制造任务中的综合基准
Ahmadreza Eslaminia, Adrian Jackson, Beitong Tian, Avi Stern, Hallie Gordon, Rajiv Malhotra, Klara Nahrstedt, Chenhui Shao
机构
*
Department of Mechanical Science and Engineering, University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校机械科学与工程系)
;
Coordinated Science Laboratory, University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校协调科学实验室)
;
Department of Mechanical and Aerospace Engineering, Rutgers University(罗格斯大学机械与航空航天工程系)
;
Department of Mechanical Engineering, University of Michigan(密歇根大学机械工程系)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.LG
机构
*
University of Connecticut(康涅狄格大学)
;
University of Massachusetts, Amherst(马萨诸塞大学阿默斯特分校)
;
School of Computing, and Augmented Intelligence, Arizona State University(亚利桑那州立大学计算与增强智能学院)
;
UMass Chan Medical School(马萨诸塞大学陈医学院)
;
University of Minnesota(明尼苏达大学)
;
Rollins School of Public Health, Emory University(埃默里大学罗林斯公共卫生学院)
;
Optum AI
;
University of Massachusetts, Lowell(马萨诸塞大学洛厄尔分校)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL
A Semi-Automated Annotation Workflow for Paediatric Histopathology Reports Using Small Language Models
一种用于儿童病理科报告的半自动化标注工作流程使用小型语言模型
Avish Vijayaraghavan, Jaskaran Singh Kawatra, Sebin Sabu, Jonny Sheldon, Will Poulett, Alex Eze, Daniel Key, John Booth, Shiren Patel, Jonny Pearson, Dan Schofield, Jonathan Hope, Pavithra Rajendran, Neil Sebire
机构
*
Imperial College London(帝国理工学院)
;
NHS England(英国国家医疗服务体系)
;
Great Ormond Street Hospital(大奥蒙德街儿童医院)
;
University College London(伦敦大学学院)
专题命中
评测与基准
:language model(title,abstract);small language model(title,abstract);large language model(abstract);分类 cs.CL
The limits of bio-molecular modeling with large language models : a cross-scale evaluation
大语言模型在生物分子建模中的局限性:跨尺度评估
Yaxin Xu, Yue Zhou, Tianyu Zhao, Fengwei An, Zhixiang Ren
机构
*
Southern University of Science and Technology(南方科技大学)
;
Pengcheng Laboratory(鹏城实验室)
;
Institute of Mechanics, Chinese Academy of Sciences(中国科学院力学研究所)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.LG