arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 431 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 431 篇

2605.14380 2026-08-10 cs.CL 版本更新 57%

VISHC at PsyDefDetect: Mitigating Data Scarcity in Psychological Defense Classification with Context-Aware Synthetic Augmentation

通过上下文感知的合成增强缓解心理防御分类中的数据稀缺

Hoang-Thuy-Duong Vu, Quoc-Cuong Pham, Huy-Hieu Pham

机构 * College of Engineering and Computer Science, VinUniversity, Hanoi, Vietnam(越南 Vin大学工程与计算机科学学院,河内,越南) VinUni-Illinois Smart Health Center, VinUniversity, Hanoi, Vietnam(越南 Vin大学与伊利诺伊大学智能健康中心,河内,越南) Center for Innovations in Health Sciences, VinUniversity, Hanoi, Vietnam(越南 Vin大学健康科学创新中心,河内,越南)

专题命中 领域大模型 :prompting(abstract);分类 cs.CL

AI总结 本文提出上下文感知合成增强框架与混合分类模型,解决心理防御机制分类中的数据稀缺问题,实验表明方法在低资源环境下取得显著提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.02491 2026-08-06 cs.AI 版本更新 57%

Long-term Measurements: Towards a Longitudinal Understanding of Human-AI Interactions

长期测量:迈向对人机交互的纵向理解

Nicole Mitchell, Dhruv Agarwal, Maty Bohacek, Remi Denton, Roma Patel

机构 * Google Research(谷歌研究院) Cornell University(康奈尔大学) Stanford University(斯坦福大学) Google DeepMind(谷歌DeepMind)

专题命中 领域大模型 :language model(abstract);分类 cs.AI

AI总结 本研究针对语言模型融入生活引发的长期人机交互风险,结合社会科学测量与NLP计算方法,提出通过长期测量建模人类行为变化,实现问题行为在线检测以缓解用户长期风险。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.12153 2026-08-06 cs.SE cs.AI 版本更新 57%

CIDR: A Large-Scale Industrial Source Code Dataset for Software Engineering Research

CIDR:一个大规模工业源代码数据集用于软件工程研究

Vladislav Savenkov

机构 * Fermatix AI

专题命中 领域大模型 :language model(abstract);分类 cs.AI

AI总结 CIDR是首个基于工业合作的大型源代码数据集,包含2440个仓库、138种语言和3.73亿行代码,支持代码智能、软件质量分析和语言模型预训练等研究。

Comments 60 pages, 13 figures, 8 appendices. Dataset access: https://fermatix.ai/#Contact. Anonymization tool: https://github.com/Fermatix/repo-sanitizer. Metadata utility: https://github.com/Fermatix/repo_metadata_cli

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.15314 2026-08-05 cs.AI 版本更新 57%

Cura 1T: Specialized Model for Agentic Healthcare

Cura 1T:用于智能医疗保健的专用模型

actAVA AI, :, Haolin Chen, Leon Qi, Steve Brown, Deon Metelski, Tao Xia, Joonyul Lee, Qixuan Wang, Kevin Riley, Frank Wang, Weiran Yao

机构 * actAVA AI(actAVA人工智能公司)

专题命中 领域大模型 :LLM(abstract);分类 cs.AI

AI总结 研究针对医疗保健中多任务需求及能力失效问题,提出通过人工门控自我进化循环训练的Cura 1T模型,以数据为中心改进模型,该模型在医疗评估套件中表现优异,在相关基准测试中保持竞争力。

Comments Model: https://actava.ai/cura; Docs: https://actava.ai/cura/docs; Github: https://github.com/actava-ai/Cura

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.18248 2026-07-31 cs.CR cs.CL 版本更新 57%

Beyond Pattern Matching: Seven Cross-Domain Techniques for Prompt Injection Detection

超越模式匹配:七种跨领域技术用于提示注入检测

Thamilvendhan Munirathinam

机构 * Independent Researcher(独立研究者) prompt-shield project(prompt-shield项目)

专题命中 领域大模型 :LLM(abstract);分类 cs.CL

AI总结 本文提出七种跨领域技术用于提示注入检测,通过引入来自不同领域的机制,如法医学语言学、材料科学疲劳分析、网络安全欺骗技术、生物信息学本地序列比对、经济学机制设计、流行病学谱信号分析和编译器理论中的污点跟踪,以改进现有的提示注入检测方法。

Comments v4 (31 pp, up from 27): adds Sec. 2.4 concurrent-work map (13 papers), Sec. 4.3 marked implemented (d034 ships in v0.7.3), Sec. 5.7 composed-stack adaptive-attack partial run, Sec. 5.8 50-doc held-out benchmark (d027 1.000->0.000 F1 in isolation, composed engine 0.815 F1), Sec. 7 architectural patterns. Repro tag: v0.7.3. Zenodo DOI 10.5281/zenodo.19644135

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.13174 2026-07-31 eess.IV cs.CV cs.LG 版本更新 57%

Scalable Drift Monitoring in Medical Imaging AI

医学影像AI中的可扩展漂移监测

Jameson Merkow, Felix J. Dorfner, Xiyu Yang, Alexander Ersoy, Giridhar Dasegowda, Mannudeep Kalra, Matthew P. Lungren, Christopher P. Bridge, Ivan Tarapov

专题命中 领域大模型 :foundation model(abstract);分类 cs.LG

AI总结 本研究针对医学影像AI的模型漂移与可靠性问题,开发了基于CheXstray框架的增强型可扩展漂移监测框架MMC+,经真实世界数据验证可有效检测数据偏移并预警性能偏差,助力AI在临床场景的应用。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.04705 2026-07-28 cs.CV cs.AI 版本更新 57%

Enhancing MedSAM with a Lightweight Box Predictor for Medical Image Segmentation

通过轻量级框预测器增强 MedSAM 用于医学图像分割

Amirhossein Movahedisefat, Amirreza Fateh, Mohammad Reza Mohammadi

机构 * School of Computer Engineering, Iran University of Science and Technology (IUST)(伊朗科学技术大学计算机工程学院)

专题命中 领域大模型 :foundation model(abstract);分类 cs.AI

AI总结 提出一种集成轻量级框预测器的 MedSAM 增强框架,通过单次点击估计边界框以提升点提示的空间引导能力,在仅增加 1.6M 参数下显著提高多模态医学图像分割的准确性和鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.02366 2026-07-28 cs.IR cs.AI 版本更新 57%

TextBridgeGNN: Pre-training Graph Neural Network for Cross-Domain Recommendation via Text-Guided Transfer

TextBridgeGNN: 通过文本引导的迁移预训练图神经网络进行跨域推荐

Yiwen Chen, Yiqing Wu, Huishi Luo, Fuzhen Zhuang, Deqing Wang, Zhao Zhang

机构 * Beihang University(北京航空航天大学) Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所) Institute of Artificial Intelligence, Beihang University(北京航空航天大学人工智能研究院)

专题命中 领域大模型 :language model(abstract);分类 cs.AI

AI总结 TextBridgeGNN通过多级图传播利用文本建立领域间关系,解决ID嵌入非转移性和异构图结构不兼容问题,提升跨域推荐性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16170 2026-07-28 cs.IR cs.AI 版本更新 57%

EGRA:Toward Enhanced Behavior Graphs and Representation Alignment for Multimodal Recommendation

EGRA:迈向用于多模态推荐的增强行为图和表示对齐

Xiaoxiong Zhang, Xin Zhou, Zhiwei Zeng, Yongjie Wang, Zhiqi Shen

机构 * College of Computing and Data Science, Nanyang Technological University, Singapore(南洋理工大学计算与数据科学学院)

专题命中 领域大模型 :prompting(abstract);分类 cs.AI

AI总结 针对多模态推荐现有方法不足,EGRA通过纳入预训练模型生成表示构建的项-项图减轻稀疏性,引入双层动态对齐加权机制提升模态-行为表示对齐,实验证明其有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.28390 2026-07-23 cs.AI 版本更新 57%

You Live More Than Once: Towards Hierarchical Skill Meta-Evolving

你活不止一次:迈向分层技能元进化

Xujun Li, Kehan Zheng, Mingyuan Zhao, Yize Geng, Jinfeng Zhou, Qi Zhu, Fei Mi, Lifeng Shang, Minlie Huang, Hongning Wang

机构 * Department of Computer Science and Technology, Tsinghua University(清华大学计算机科学与技术系) Huawei Foundation Model Department(华为基础模型部门)

专题命中 领域大模型 :LLM(abstract_cn);分类 cs.AI

AI总结 本文提出HiSME,一种轻量级分层技能元进化方法,通过从智能体任务执行轨迹中学习元技能,联合优化技能和技能进化策略,以持续提升部署的智能体系统在不同下游场景中的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.17719 2026-07-22 cs.AI cs.MA 版本更新 57%

SR-Agent: An Experience-Driven Agentic Framework for Post-Ranking Strategy Refinement in E-Commerce Recommendation

SR-Agent:一种用于电子商务推荐中排序后策略优化的经验驱动智能框架

Hanchen Yang, Kaiwen Yang, Junpeng Zhuang, Yang He, Keting Cen, Bochao Liu, Zhongbo Sun, An Liu, Zhongteng Han, Chenyi Lei

机构 * Kuaishou Technology(快手科技)

专题命中 领域大模型 :LLM(abstract);分类 cs.AI

AI总结 电商推荐系统中,后排序策略因环境变化需优化,以往方式存在问题。本文提出SR-Agent框架,统一用户模拟、分析和策略优化组件,经快手平台测试,能提升订单量、浏览深度和点击类别多样性,还缩短优化周期、降低成本。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.09757 2026-07-21 cs.CV cs.AI 版本更新 57%

MedLVR: Latent Visual Reasoning for Reliable Medical Visual Question Answering

MedLVR: 基于潜在视觉推理的可靠医学视觉问答

Suyang Xi, Songtao Hu, Yuxiang Lai, Wangyun Dan, Yaqi Liu, Shansong Wang, Xiaofeng Yang

机构 * Department of Radiation Oncology and Winship Cancer Institute, Emory University School of Medicine(埃默里大学医学院放射肿瘤学系与温希普癌症研究所) Department of Biostatistics and Bioinformatics, Emory University(埃默里大学生物统计学与生物信息学系)

专题命中 领域大模型 :language model(abstract);分类 cs.AI

AI总结 MedLVR通过引入显式视觉证据状态,在自回归解码中插入短时latent推理段,提升医学视觉问答的可靠性与准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01086 2026-07-17 cs.AI cs.CR cs.DB cs.DC cs.SE 版本更新 57%

MedBeads: An AI-Native Clinical Context Graph Built from Immutable Beads and Reconstructable Clinical Links

MedBeads:面向可信医疗AI的智能体原生不可变数据基底

Takahito Nakajima

机构 * Diagnostic Imaging and Interventional Radiology, Institute of Medicine, University of Tsukuba(东京大学医学研究院诊断影像与介入放射学部) Center for Cyber Medicine Research, University of Tsukuba(东京大学计算机医学研究中心)

专题命中 领域大模型 :LLM(abstract_cn);分类 cs.AI

AI总结 针对医疗AI中电子病历与智能体间的上下文不匹配问题,提出基于Merkle有向无环图的不可变数据架构MedBeads,通过确定性图遍历替代概率检索,实现可审计、防篡改的临床上下文提供。

Comments 23 pages, 5 figures, 3 tables. Reference implementation and reproducible Docker demo available at https://github.com/medbeads/medbeads

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.12441 2026-07-16 cs.CL 版本更新 57%

WikiSTAR: A System for Shedding Light on the Hidden History of Scientific Wikipedia Articles

WikiSTAR:一个揭示科学维基百科文章隐藏历史的系统

Omer Ehrlich, Nitzan Barzilay, Rona Aviram, Tom Hope

机构 * The Hebrew University of Jerusalem(耶路撒冷希伯来大学) Ben-Gurion University of the Negev(内盖夫本-古里安大学) Allen Institute for AI (Ai2)(艾伦人工智能研究所)

专题命中 领域大模型 :LLM(abstract);分类 cs.CL

AI总结 WikiSTAR系统利用大语言模型分类器标记编辑类型,通过交互式视图追溯科学维基百科文章修订历史,揭示知识发展,经用户研究验证其能发现新模式、问题并实现新分析,还发布了系统、代码和基准。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.14999 2026-07-16 cs.LG 版本更新 57%

Unlocking Latent Dimensions: Exploring Representations of Large-Scale X-ray Scattering Data using Variational Autoencoders

解锁潜在维度:使用变分自编码器探索大规模X射线散射数据的表示

Monika Choudhary, Xiaoya Chong, Runbo Jiang, Wiebke Koepp, Petrus H. Zwart, Damon English, Gregory M. Su, Eric Schaible, Chenhui Zhu, Mostafa Nassr, Noah P. Wamble, Kelvin Kam-Yun Li, Jonathan M. Chan, Jose Carlos Diaz, Cameron McKay, Lynn Katz, Benny Freeman, Guillaume Freychet, Yevgen Matviychuk, Eliot Gann, Daniel B. Allan, Benedikt Sochor, Frank Schluenzen, Stephan V. Roth, Ethan J. Crumlin, Dylan McReynolds, Tanny Chavez, Alexander Hexemer

机构 * Advanced Light Source, Lawrence Berkeley National Laboratory(劳伦斯伯克利国家实验室先进光源) Center for Advanced Mathematics for Energy Research Applications, Lawrence Berkeley National Laboratory(劳伦斯伯克利国家实验室能源研究应用高级数学中心) Molecular Biophysics & Integrated Bioimaging Division, Lawrence Berkeley National Laboratory(劳伦斯伯克利国家实验室分子生物物理学与综合生物成像部) Berkeley Synchrotron Infrared Structural Biology program, Lawrence Berkeley National Laboratory(劳伦斯伯克利国家实验室伯克利同步辐射红外结构生物学项目) Materials Sciences Division, Lawrence Berkeley National Laboratory(劳伦斯伯克利国家实验室材料科学部) McKetta Department of Chemical Engineering, University of Texas(德克萨斯大学麦凯塔化学工程系)

专题命中 领域大模型 :foundation model(abstract);分类 cs.LG

AI总结 针对X射线散射数据离线探索和实时分析两大挑战,训练领域特定注意力卷积变分自编码器(C-VAE),学习低维表示以捕捉结构变化,并集成到MLExchange平台的Latent Space Explorer中,支持交互式结构探索。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.14473 2026-07-16 cs.CL 版本更新 57%

AI Can Learn Scientific Taste

AI 可以学习科学品味

Jingqi Tong, Mingzhe Li, Hangcheng Li, Yongzhuo Yang, Yurong Mou, Weijie Ma, Zhiheng Xi, Hongji Chen, Xiaoran Liu, Qinyuan Cheng, Ming Zhang, Qiguang Chen, Weifeng Ge, Qipeng Guo, Tianlei Ying, Tianxiang Sun, Yining Zheng, Xinchi Chen, Jun Zhao, Ning Ding, Xuanjing Huang, Yu-Gang Jiang, Xipeng Qiu

专题命中 领域大模型 :LLM(abstract);分类 cs.CL

AI总结 本文提出RLCF框架,通过社区反馈学习科学品味,使AI能提出高潜力研究想法,实验显示其优于现有模型并具备泛化能力。

Comments 46 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.15748 2026-07-15 cs.LG cs.CV 版本更新 57%

Visual Species Recognition with Large Multimodal Models as Post-Hoc Correctors

以大型多模态模型作为事后校正器的视觉物种识别

Tian Liu, Anwesha Basu, James Caverlee, Shu Kong

机构 * Texas A&M University(德克萨斯A&M大学) University of Macau(澳门大学) Institute of Collaborative Innovation(协同创新研究院)

专题命中 领域大模型 :prompting(abstract);分类 cs.LG

AI总结 研究视觉物种识别问题,比较少样本学习专家模型与大型多模态模型后发现后者虽有不足但有互补优势,进而提出事后校正框架,利用多模态提示策略提升少样本学习专家模型准确率,该框架具有通用性和有效性。

Comments website and code: https://tian1327.github.io/POC

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00972 2026-07-14 physics.data-an cs.AI cs.CV cs.IR 版本更新 57%

Toward a Scientific Discovery Engine for Weather and Climate Data: A Visual Analytics Workbench for Embedding-Based Exploration

迈向面向天气和气候数据的科学发现引擎:一种用于基于嵌入探索的可视化分析工作台

Nihanth W. Cherukuru, Matt Rehme, Kirsten J. Mayer, David John Gagne, John Schreck, John Clyne, Charlie Becker

机构 * NSF National Center for Atmospheric Research(国家大气研究中心)

专题命中 领域大模型 :foundation model(abstract);分类 cs.AI

AI总结 本文提出了一种开源可视化分析工作台,用于探索基于嵌入的天气和气候数据,通过链接嵌入实验与源数据、元数据、空间上下文和模型配置,使潜在空间结果可追溯到物理过程,支持科学发现流程。

Comments 7 pages, 5 figures, Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.13330 2026-07-14 cs.CL 版本更新 57%

RegCheck: A tool for structured comparisons between study registrations and papers

RegCheck:一个用于研究注册与论文之间结构化比较的工具

Jamie Cummins, Beth Clarke, Ian Hussey, Malte Elson

机构 * Bennett Institute of Applied Data Science, University of Oxford(贝内特应用数据科学研究所,牛津大学) Institute of Psychology, University of Bern(心理学研究所,伯恩大学)

专题命中 领域大模型 :LLM(abstract);分类 cs.CL

AI总结 介绍RegCheck工具,用于比较研究注册与论文。利用人工智能,模块化设计,让用户决定比较特征并提供相关文本,生成可共享报告,适用于多领域多格式,有望成为可重复科学的可扩展基础设施。

Comments 26 pages, 4 figures, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.21414 2026-07-10 cs.CV cs.LG 版本更新 57%

A Tool Bottleneck Framework for Clinically-Informed and Interpretable Medical Image Understanding

用于临床信息丰富且可解释的医学图像理解的工具瓶颈框架

Christina Liu, Alan Q. Wang, Joy Hsu, Jiajun Wu, Ehsan Adeli

机构 * California Institute of Technology(加州理工学院) Stanford University(斯坦福大学)

专题命中 领域大模型 :language model(abstract);分类 cs.LG

AI总结 针对医学图像理解中工具组合难的问题,提出工具瓶颈框架(TBF),利用工具瓶颈模型(TBM)组合VLM选择的工具,通过神经网络计算融合工具输出,在组织病理学和皮肤病学任务中表现出色,提升医学图像理解且使预测更具可解释性。

Journal ref Proceedings of the 9th International Conference on Medical Imaging with Deep Learning, Proceedings of Machine Learning Research 315 (2026) 2958-2986

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.26266 2026-07-01 cs.AI cs.CV 版本更新 57%

GUIDE: Resolving Domain Bias in GUI Agents through Real-Time Web Video Retrieval and Plug-and-Play Annotation

GUIDE:通过实时网络视频检索和即插即用标注解决GUI代理的领域偏见

Rui Xie, Zhi Gao, Chenrui Shi, Zirui Shang, Lu Chen, Qing Li

机构 * Shanghai Jiao Tong University(上海交通大学) State Key Laboratory for General Artificial Intelligence, BIGAI(通用人工智能国家重点实验室,北京通用人工智能研究院) Beijing Institute of Technology(北京理工大学)

专题命中 领域大模型 :language model(abstract);分类 cs.AI

AI总结 GUIDE通过实时网络视频检索和即插即用标注框架,解决GUI代理的领域偏见问题,通过视频语义分析和自动化标注流程提升代理对特定应用的操作流程和UI布局的理解,实验表明其在多代理系统和单模型代理中均能提升性能。

Comments Accepted to ECCV 2026. 30 pages: 15-page main paper followed by supplementary material as an appendix (Sections A-F). Project page: https://sharryXR.github.io/GUIDE/

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.26136 2026-06-26 eess.AS cs.CL 版本更新 57%

One Voice, Many Tongues: Cross-Lingual Voice Cloning for Scientific Speech

一个声音,多种语言:面向科学演讲的跨语言语音克隆

Amanuel Gizachew Abebe, Yasmin Moslem

机构 * Shaggar Institute of Technology(谢加尔技术学院) Trinity College Dublin(都柏林三一学院)

专题命中 领域大模型 :foundation model(abstract);分类 cs.CL

AI总结 本文针对跨语言科学演讲语音生成中的声音身份保持问题,提出基于OmniVoice基础模型的语音克隆系统,通过多模型集成蒸馏提升生成语音的可懂度和说话人相似性。

Comments In Proceedings of the 23rd International Conference on Spoken Language Translation (IWSLT 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.04579 2026-06-23 cs.AI 版本更新 57%

SCI-PRM: A Tool Aware Process Reward Model for Scientific Reasoning Verification

SCI-PRM:用于科学推理验证的工具感知过程奖励模型

Xiangyu Zhao, Henry Hengyuan Zhao, Yiheng Wang, Wanghan Xu, Yuhao Zhou, Qinglong Cao, Zhiwang Zhou, Lei Bai, Wenlong Zhang, Xiao-Ming Wu

机构 * The Hong Kong Polytechnic University(香港理工大学) Shanghai AI Lab(上海人工智能实验室) National University of Singapore(新加坡国立大学) Shanghai Jiao Tong University(上海交通大学) Sichuan University(四川大学) Tongji University(同济大学)

专题命中 领域大模型 :foundation model(abstract);分类 cs.AI

AI总结 针对科学推理中工具使用和事实一致性问题,提出Sci-PRM模型,通过构建包含工具链轨迹的数据集SCIPRM70K并训练过程奖励模型,在测试时扩展和强化学习中提供细粒度监督,提升基础模型性能。

Comments Accepted by KDD 2026 AI4Science Track

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.21563 2026-06-23 cs.LG 版本更新 57%

Embedding-Based Federated Learning with Runtime Governance for Iron Deficiency Prediction

基于运行时治理的嵌入式联邦学习用于缺铁预测

Fan Zhang, Simon Deltadahl, Majid Lotfian Delouee, Daniel Kreuter, Joseph Taylor, Allerdien Visser, BloodCounts Consortium, James H. F. Rudd, Nicholas S. Gleadall, Suthesh Sivapalaratnam, Folkert Asselbergs, Martijn C. Schut, Michael Roberts

机构 * Theoretical Physics University of Cambridge Cambridge, UK Translational AI Laboratory, Dept. of Laboratory Medicine Amsterdam UMC Amsterdam, The Netherlands Precision Health University Research Institute Queen Mary Univ. of London London, UK Department of Medicine University of Cambridge Cambridge Biomedical Campus Cambridge, UK Transplant Cambridge Biomedical Campus Cambridge, UK Dept. of Cardiology Amsterdam Cardiovascular Sciences Amsterdam UMC Amsterdam, The Netherlands

专题命中 领域大模型 :foundation model(abstract);分类 cs.LG

AI总结 本文提出了一种基于嵌入的联邦学习框架,用于从常规全血计数数据中预测缺铁,并在两个临床环境中部署,展示了个性化聚合方法在处理不同样本量和任务相关性时的优越性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.16625 2026-06-23 cs.AI 版本更新 57%

MedBayes-Lite: A Clinical Uncertainty Governance Layer for Risk-Aware Medical Decision Support

MedBayes-Lite: 用于风险感知医疗决策支持的临床不确定性治理层

Elias Hossain, Md Mehedi Hasan Nipu, Maleeha Sheikh, Tasfia Nuzhat, Rajib Rana, Subash Neupane, Björn W. Schuller, Niloofar Yousefi

机构 * College of Engineering and Computer Science, University of Central Florida(中央佛罗里达大学工程与计算机科学学院) Department of Computer Science and Engineering, North South University(北方南大学计算机科学与工程系) Department of Electrical and Computer Engineering, Purdue University Fort Wayne(普渡大学福克斯堡分校电气与计算机工程系) School of Mathematics, Physics and Computing, University of Southern Queensland(南方昆士兰大学数学、物理与计算学院) Meharry Medical College(梅哈里医学学院) CHI – Chair of Health Informatics, Technical University of Munich (TUM)(慕尼黑技术大学健康信息学系) GLAM – Group on Language, Audio, & Music, Imperial College London(伦敦帝国理工学院语言、音频与音乐组)

专题命中 领域大模型 :language model(abstract);分类 cs.AI

AI总结 提出无需重训练的MedBayes-Lite层,结合MC Dropout、校准和置信度引导的弃权,减少临床问答中高置信度错误,将校准误差降低0.23-0.33,高严重性错误降至近零。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.12567 2026-06-18 cs.CV cs.AI 版本更新 57%

Pyramid Self-Contrastive Learning for Single-shot Test-time Ultrasound Image Denoising

金字塔自对比学习框架用于测试时超声图像去噪

Jiajing Zhang, Bingze Dai, Xi Zhang, Yue Xu, Wei-Ning Lee

机构 * Department of Electrical and Computer Engineering, The University of Hong Kong(香港大学电子与计算机工程系) Department of Biomedical Engineering, Duke University(达特茅斯大学生物医学工程系)

专题命中 领域大模型 :pretraining(abstract);分类 cs.AI

AI总结 本文提出一种纯测试时训练框架,用于单次超声图像去噪,应用于合成孔径超声,通过自对比学习分离解剖相似性和噪声随机性,提升去噪效果和结构细节。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.18313 2026-06-16 cs.CV cs.AI 版本更新 57%

Wasserstein Equilibrium Decoding for Reliable Medical Visual Question Answering

Wasserstein均衡解码用于可靠的医疗视觉问答

Luca Hagen, Johanna P. Müller, Weitong Zhang, Mengyun Qiao, Bernhard Kainz

机构 * Friedrich-Alexander University Erlangen-Nürnberg(弗里德里希-亚历山大厄林根-纽伦堡大学) Imperial College London(伦敦帝国理工学院) University College London(伦敦大学学院)

专题命中 领域大模型 :language model(abstract);分类 cs.AI

AI总结 本文提出了一种基于Wasserstein距离的均衡解码方法,用于改进医疗视觉问答系统,通过语义感知的停止准则提高解码效率和准确性,同时在VQA-RAD和PathVQA数据集上实现了显著的性能提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.12514 2026-06-16 cs.CV cs.LG 版本更新 57%

CT-VDETR: Semi-supervised 3D Trauma Detection in Computed Tomography (CT) scans using Dense Vertex Relative Position Encoding

CT-VDETR:使用密集顶点相对位置编码的CT扫描半监督3D创伤检测

Shivam Chaudhary, Sheethal Bhat, Andreas Maier

机构 * University of Freiburg(弗赖堡大学)

专题命中 领域大模型 :pretraining(abstract);分类 cs.LG

AI总结 提出CT-VDETR框架,结合自监督预训练和半监督transformer检测,在仅78个标注体数据上实现31.33% mAP@0.50,比纯监督方法提升1.53倍。

Comments v2: Updated results with corrected dataset split. Revised Table 1 (mAP@0.50: 31.33% SSL vs 20.45% baseline, 1.53x improvement; mAP@0.75: 30.95% vs 10.45%, 2.96x improvement). Updated validation curves showing stable convergence. No methodology changes. 7 pages, 4 figures, 2 tables. Code: https://github.com/shivasmic/3d-trauma-detection-ssl

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.18784 2026-06-15 q-fin.RM cs.AI cs.CR cs.CY econ.GN q-fin.EC 版本更新 57%

The Insurability Frontier of AI Risk: Mapping Threats to Affirmative Coverage, Silent Exposures, and Exclusions

AI风险的可保险边界:将威胁映射到积极保险、沉默暴露和排除

Alex Leung, Rex Zhang, Ervin Ling, Kentaroh Toyoda, SiewMei Loh

机构 * Munich Re(慕尼黑再保险) Armilla Tokio Marine Kiln(东京海上日赤保险) CFC Apollo ibott Coalition

专题命中 领域大模型 :foundation model(abstract);分类 cs.AI

AI总结 本文研究了AI风险在商业保险中的可保险性边界,通过分析55类AI威胁与26种保险产品和排除制度,揭示了四个层次的可保险性前沿:积极保险的风险、沉默AI暴露、主动排除的风险以及传统私人保险结构之外的风险。

Comments Version 2

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.04602 2026-06-12 cs.AI 版本更新 57%

Parthenon Law: A Self-Evolving Legal-Agent Framework

Parthenon Law: 一种自我进化的法律智能体框架

Hejia Geng, Leo Liu

机构 * tapntell.ai

专题命中 领域大模型 :LLM(abstract);分类 cs.AI

AI总结 本文提出Parthenon框架,通过分解模型、工具、知识等组件并引入反泄漏学习循环,使法律领域的大语言模型智能体能够从经验中自我进化,显著提升法律事务处理性能。

详情

展开后加载摘要…

URL PDF HTML 收藏