arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-06-08 至 2026-06-08 共收录 25 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 25 篇

2606.06694 2026-06-08 cs.LG cs.AI cs.CY 新提交 92%

The Geography of Algorithmic Judgment: LLM Intermediaries, Place Identity, and Racial Steering in Housing Search

算法判断的地理:LLM中介、地方身份与住房搜索中的种族引导

Hana Samad, Trung Lam, Christoph Mügge-Durum, Michael Akinwumi

机构 * National Fair Housing Institute(国家公平住房研究所)

专题命中 领域大模型 :LLM(title,title_cn);large language model(abstract);language model(abstract);prompting(abstract)

AI总结 通过行为审计七种LLM在四个美国城市的住房推荐,发现种族引导是模型解释性许可的涌现行为,而非静态属性,且城市并非中性测试单元。

Comments 13 pages with supplemental tables and figures, AIES '26 Submission

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.07237 2026-06-08 cs.CL cs.AI cs.LG 新提交 92%

When Large Language Models Fail in Healthcare: Evaluating Sensitivity to Prompt Variations

当大型语言模型在医疗保健中失败:评估对提示变化的敏感性

Mahdi Alkaeed

机构 * Department of Computer Science and Engineering, Doha, Qatar(计算机科学与工程系,多哈,卡塔尔)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(summary_cn,abstract_cn);分类 cs.CL、cs.AI、cs.LG

AI总结 本研究系统分析了通用和医学专用LLM对提示扰动的敏感性,发现即使是微小的措辞变化也可能改变临床建议,对抗性提示可能引发有害输出,表明这些模型在临床应用中不可靠。

Comments 12 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.07342 2026-06-08 cs.CL cs.NE 新提交 92%

LLM-Guided Evolution for Medical Decision Pipelines

LLM引导的医疗决策流程进化

Ivan Sviridov, Artem Oskin, Ivan Panin, Iaroslav Bespalov, Dmitry Dylov, Ivan Oseledets, Aleksandr Nesterov

机构 * Sber AI Lab(Sber AI实验室) AIRI

专题命中 领域大模型 :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 提出LLM引导的MAP-Elites进化方法,无需微调即可优化医疗决策流程,在分诊、咨询和图像分类任务中超越手工设计基线。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.10896 2026-06-08 cs.CL 版本更新 90%

DialDefer: A Framework for Detecting and Mitigating LLM Dialogic Deference

DialDefer: 检测和缓解LLM对话性遵从的框架

Parisa Rabbani, Priyam Sahoo, Ruben Mathew, Aishee Mondal, Harshita Ketharaman, Nimet Beyza Bozdag, Dilek Hakkani-Tür

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

专题命中 领域大模型 :LLM(title,title_cn);分类 cs.CL

AI总结 提出DialDefer框架,通过对话性遵从分数检测和缓解LLM在对话评估中因提问框架导致的判断偏移,发现框架效应显著但准确率稳定,且模型对人类与AI的不同归因产生最大偏移。

Comments 10 pages main content, 7 figures, 35 pages total with appendix

Journal ref ACL 2026 - Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.06823 2026-06-08 cs.LG cs.AI q-fin.ST 新提交 87%

PandaAI: A Practical Agent CQ2 for Neuro-symbolic Data Analysis And Integrated Decision-Making in Quantitative Finance

PandaAI: 一种用于量化金融中神经符号数据分析与集成决策的实用智能体CQ2

Yuqi Li, Siyuan Liu, Bingjun Liu

机构 * Panda AI

专题命中 领域大模型 :LLM(summary_cn,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 针对金融数据低信噪比和非平稳性,提出PandaAI,一种结合市场状态建模与约束alpha生成的闭环神经符号LLM智能体,通过领域微调和模块化架构实现风险感知决策,在沪深300数据上Rank IC提升18.2%,最大回撤降低25.7%。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.22327 2026-06-08 cs.IR cs.AI cs.DL 版本更新 86%

Evaluating AI-based Scientific Knowledge Synthesis with Epidemiological Systematic Reviews

基于流行病学系统评价评估AI科学知识综合

Shreyansh Padarha, Ryan Othniel Kearns, Tristan Naidoo, Lingyi Yang, Łukasz Borchmann, Piotr BŁaszczyk, Christian Morgenstern, Ruth McCabe, Sangeeta Bhatia, Philip H. Torr, Jakob Foerster, Scott A. Hale, Thomas Rawson, Anne Cori, Elizaveta Semenova, Adam Mahdi

机构 * University of Oxford(牛津大学) Imperial College London(伦敦帝国理工学院) University of Nottingham(诺丁汉大学) Snowflake AI Research(Snowflake人工智能研究) Independent(独立)

专题命中 领域大模型 :LLM(summary_cn,abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 提出AgentSLR评估框架,包含自动化工作流和专家标注数据集,测试LLM在流行病学系统评价各阶段能力,发现无模型全面领先,结构化提取是主要瓶颈。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.06760 2026-06-08 cs.CV 新提交 85%

MedSIGHT: Towards Grounded Visual Comprehension in Medical Large Vision-Language Models

MedSIGHT:迈向医学大型视觉语言模型中的基础视觉理解

Aofei Chang, Le Huang, Alex James Boyd, Parminder Bhatia, Taha Kass-Hout, Fenglong Ma, Cao Xiao

机构 * University of California, Berkeley(加州大学伯克利分校)

专题命中 领域大模型 :language model(title,abstract);LLM(abstract,abstract_cn)

AI总结 提出MedSIGHT框架,通过区域感知器、医学区域码本和渐进训练策略,统一医学视觉语言模型的语义理解和像素级分割,在72K数据上达到多模态理解与分割的SOTA。

Comments Accepted at ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25522 2026-06-08 cs.AI 版本更新 84%

Understanding Generative Recommendation with Semantic IDs from a Model-scaling View

从模型扩展视角理解基于语义ID的生成式推荐

Jingzhe Liu, Liam Collins, Jiliang Tang, Tong Zhao, Neil Shah, Clark Mingxuan Ju

机构 * Michigan State University(密歇根州立大学) Snap Inc.(Snap公司)

专题命中 领域大模型 :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);foundation model(abstract)

AI总结 揭示基于语义ID的生成式推荐在模型扩展时存在性能瓶颈,发现直接使用大语言模型作为推荐器具有更好的扩展性,性能提升可达20%。

Comments Accepted by KDD 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.23292 2026-06-08 cs.AI cs.LG 版本更新 84%

Agentic Physical AI toward a Domain-Specific Foundation Model for Energy Systems: A Case Study on Nuclear Reactor Control

面向能源系统的领域特定基础模型的具身物理人工智能:以核反应堆控制为例

Yoon Pyo Lee, Samrendra Roy, Kazuma Kobayashi, Sajedul Talukder, Diab Abueidda, Seid Koric, Souvik Chakraborty, Syed Bahauddin Alam

机构 * The Grainger College of Engineering, Nuclear, Plasma & Radiological Engineering, University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校格雷格学院工程学院、核等工程学院) Department of Nuclear Engineering, Hanyang University(汉阳大学核工程系) University of Texas - El Paso(德克萨斯大学埃尔帕索分校) National Center for Supercomputing Applications(国家超级计算应用中心) Department of Applied Mechanics, Indian Institute of Technology Delhi(印度德里理工学院应用力学系) Yardi School of Artificial Intelligence, Indian Institute of Technology Delhi(印度德里理工学院亚里人工智能学院)

专题命中 领域大模型 :foundation model(title,abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本研究提出通过紧凑语言模型作为具身物理人工智能,利用基于物理模拟器验证的策略优化替代感知推理,在核反应堆控制任务中实现领域特定基础模型,并展示了规模扩展带来的可靠性提升和策略集中化行为。

Comments Accepted for publication in npj Artificial Intelligence

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06915 2026-06-08 cs.CL cs.LG 版本更新 82%

A Dynamic Self-Evolving Extraction System

一种动态自演化抽取系统

Moin Amin-Naseri, Hannah Kim, Estevam Hruschka

机构 * Megagon Labs(Megagon实验室)

专题命中 领域大模型 :LLM(summary_cn,abstract);分类 cs.CL、cs.LG

AI总结 提出DySECT系统,通过LLM抽取三元组构建知识库,结合概率知识和图推理丰富知识,再反馈优化抽取器,形成闭环持续提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04982 2026-06-08 cs.CY cs.AI cs.HC 版本更新 81%

Training for Technology: Adoption and Productive Use of Generative AI in Legal Analysis

技术培训:法律分析中生成式人工智能的采纳与生产性使用

Benjamin M. Chen, Hong Bao

机构 * University of Hong Kong Faculty of Law(香港大学法学院)

专题命中 领域大模型 :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 通过随机实验发现,未经培训的法学学生使用大语言模型反而降低表现,而简短培训能显著提升采纳率和成绩,表明生成式AI的生产力需要培训支持。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.06749 2026-06-08 q-bio.QM 新提交 80%

Deterministic access to global viral sequence data enables robust agentic scientific discovery

确定性访问全球病毒序列数据实现稳健的自主科学发现

Ferdous Nasri, Sarah Gurev, Patrick Varilly, Krithik Ramesh, Nuala A. O'Leary, Jonah Cool, Bernhard Y. Renard, Pardis C. Sabeti, Laura Luebbert

专题命中 领域大模型 :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract)

AI总结 针对基于大语言模型的科学代理在病毒数据检索中的高错误率问题,提出确定性查询框架gget virus,通过形式化NCBI Virus过滤流程、元数据约束和结构化记录检索,将检索准确率提升至90%以上,并减少98%数据传输。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17601 2026-06-08 cs.IR cs.SI 80%

Why They Link: An Intent Taxonomy for Including Hyperlinks in Social Posts

为何他们链接:一种包含超链接在社交帖子中的意图分类

Fangping Lan, Abdullah Aljebreen, Eduard C. Dragut

专题命中 领域大模型 :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract)

AI总结 本文提出一种意图分类,通过混合方法分析用户对超链接的感知意图,揭示广告、论证和分享是最常见的意图,并在微博客检索任务中展示其应用价值。

Comments 10 pages including references, 5 figures,

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.06534 2026-06-08 eess.IV cs.AI 新提交 79%

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models

基于视觉基础模型的注意力一致纵向医学视觉问答

Jialin Wu, Qianru Zhang, Georges El Fakhri, Xiaofeng Liu

机构 * University of California, San Diego(加州大学圣地亚哥分校) Yale Biomedical Imaging Institute(耶鲁大学生物医学成像研究所)

专题命中 领域大模型 :foundation model(title,abstract);分类 cs.AI

AI总结 提出一种注意力引导的编码器-解码器框架,通过轻量级配准和自适应掩码生成,结合辅助损失函数,实现胸部X光片的纵向医学视觉问答,在Medical-Diff-VQA基准上取得优异性能。

Comments Accepted to CVPR 2026 Workshop PHAROS-AIF-MIH

Journal ref Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) Workshops, 2026, pp. 6448-6458

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.04550 2026-06-08 cs.IR cs.AI cs.SI 交叉投稿 77%

Trading Engagement for Sustainability: Carbon-Aware Re-ranking for E-commerce Recommendations

用参与度换取可持续性:面向电子商务推荐中碳感知的重排序

Noah Lund Syrdal, Anders Vestrum, Jorgen Bergh

机构 * University of California, Berkeley(加州大学伯克利分校)

专题命中 领域大模型 :LLM(abstract,abstract_cn);prompting(abstract);分类 cs.AI

AI总结 本文提出一种碳感知重排序策略,通过检索增强的碳足迹估计管道推断缺失的产品碳足迹标签,并在三个推荐模型上权衡用户参与度与碳排放,实现可持续推荐。

Comments 23 pages, 30 figures. Code available at https://github.com/andersvestrum/carbon-aware-recsys

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17004 2026-06-08 cs.MA cs.AI 版本更新 77%

ReclAIm: A Multi-Agent Framework for Monitoring and Correcting Performance Decline in Medical Imaging AI

ReclAIm:用于监测和纠正医学影像AI性能下降的多智能体框架

Eleftherios Tzanis, Michail E. Klontzas

机构 * Artificial Intelligence and Translational Imaging (ATI) Lab, Department of Radiology, School of Medicine, University of Crete(人工智能与转化成像实验室,放射科,医学院,希腊克里特大学) Computational Biomedicine Laboratory, Institute of Computer Science Foundation for Research and Technology Hellas (ICS - FORTH), Heraklion, Crete, Greece(计算生物医学实验室,希腊基础研究与技术院计算机科学研究所(ICS - FORTH),克里特,希腊) Division of Radiology, Department of Clinical Science, Intervention and Technology (CLINTEC), Karolinska Institute, Huddinge, Sweden(放射科,临床科学、干预与技术部(CLINTEC),卡罗林斯卡研究所,瑞典Huddinge)

专题命中 领域大模型 :large language model(abstract);language model(abstract);prompting(abstract);分类 cs.AI

AI总结 提出基于大语言模型的多智能体框架ReclAIm,通过自然语言交互自动监测医学图像分类模型性能下降并触发微调,采用数据增强、类别不平衡处理和参数锚定正则化策略,在多个数据集上验证了有效性。

Comments Published in Radiology: Artificial Intelligence (https://doi.org/10.1148/ryai.250923)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.24481 2026-06-08 cs.AI cs.CL cs.LG 版本更新 75%

Multi-Agent Reasoning with Consistency Verification Improves Uncertainty Calibration in Medical MCQA

基于一致性验证的多智能体推理改进医学多项选择题问答中的不确定性校准

John Ray B. Martinez

机构 * Department of Data Science and Analytics(数据科学与分析系)

专题命中 领域大模型 :LLM(abstract,abstract_cn);分类 cs.CL、cs.AI、cs.LG

AI总结 提出多智能体框架,结合领域专家智能体与两阶段验证及S分数加权融合,在医学MCQA中显著降低校准误差并提升判别能力。

Comments 20 pages, 6 figures. Preprint under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.07113 2026-06-08 cs.AI 新提交 70%

Beyond Post-hoc Explanation: Toward Glassbox AI via Probabilistic Mediation

超越事后解释:通过概率中介迈向玻璃箱AI

Manuele Leonelli

机构 * Manuele Leonelli(曼努埃尔·莱奥内利)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 针对大语言模型在关键领域的不透明性,提出玻璃箱框架,利用贝叶斯网络作为事前中介层,实现可审计推理、不确定性量化和可争议输出。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07380 2026-06-08 cs.CL 版本更新 70%

Mining Useful General Data for Low-Resource Domain Adaptation

挖掘低资源领域适应的有用通用数据

Pingjie Wang, Hongcheng Liu, Yusheng Liao, Ziqing Fan, Yaxin Du, Shuo Tang, Yanfeng Wang, Yu Wang

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 针对低资源领域数据稀缺问题,提出NTK-Selector方法,利用神经正切核从通用数据中筛选有用样本,显著提升领域适应效果。

Comments 39 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.05232 2026-06-08 cs.AI 版本更新 70%

ChemQuests: A Curated Chemistry Question-Answer Database Extracted from ChemRxiv papers

ChemQuests: 从ChemRxiv论文中提取的精选化学问答数据库

Mahmoud Amiri, Thomas Bocklitz

机构 * University of California, Berkeley(加州大学伯克利分校) Stanford University(斯坦福大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 提出ChemQuests数据集,包含从155篇ChemRxiv论文中提取的952个高质量问答对,覆盖17个化学子领域,用于化学NLP研究。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.05949 2026-06-08 cs.CV 版本更新 67%

Faithful, Enriched, and Precise: Benchmarking Natural-Science Illustration Generation by T2I models

忠实、丰富且精确:T2I模型在自然科学插图生成中的基准测试

Yifan Chang, Jiaxin Ai, Jianwen Sun, Yuandong Pu, Siqi Luo, Liangliang Zhao, Yuchen Ren, Minghao Liu, Yunfei Yu, Yu Qiao, Kaipeng Zhang, Yihao Liu

机构 * Shanghai Innovation Institute(上海创新研究院) Shanghai AI Laboratory(上海人工智能实验室) University of Science and Technology of China(中国科学技术大学) Wuhan University(武汉大学) Nankai University(南开大学) Shanghai Jiao Tong University(上海交通大学) Fudan University(复旦大学) ZODA Alaya Studio(Alaya工作室)

专题命中 领域大模型 :large language model(abstract);language model(abstract)

AI总结 提出FEPBench基准,通过细粒度原子集标注和三维评估(指令忠实性、推理丰富性、语义精确性)系统评估T2I模型在自然科学插图生成中的表现,发现即使最先进的闭源模型仍存在文本渲染瓶颈、推理丰富性有限以及生成丰富性与精确性难以平衡的问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00909 2026-06-08 cs.IR 版本更新 67%

HiPS: Hierarchical PDF Segmentation of Doctrinal Legal Books

HiPS: doctrinal法律书籍的分层PDF分割

Sabine Wehnert, Harikrishnan Changaramkulath, Ivan Habernal

专题命中 领域大模型 :LLM(abstract,abstract_cn)

AI总结 本文提出HiPS,用于分层PDF分割doctrinal法律书籍,提供了一个包含49本开放获取法律书籍的黄金标准基准,包含9812个手动编纂的标题、层级和页面锚点,同时引入了基于目录和无目录的分割管道,以提高标题检测、层级重建和边界分配的准确性。

Comments 11 pages, 9 figures. Accepted as a demo paper at ICAIL 2026. This arXiv version includes an appendix, new results, bug fixes, and presentation improvements beyond the earlier preprint; consequently, some reported numbers differ

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.15888 2026-06-08 cs.LG cs.AI 版本更新 62%

CHoE: Cross-Domain Heterogeneous Graph Prompt Learning via Structure-Conditioned Experts

CHoE: 基于结构条件专家的跨域异构图提示学习

Peiyuan Li, Yongqi Huang, Jitao Zhao, Dongxiao He, Di Jin, Weixiong Zhang

机构 * School of Computer Science and Technology, Tianjin University(天津大学计算机科学与技术学院) Department of Health Technology and Informatics, and Department of Data Science and Artificial Intelligence, The Hong Kong Polytechnic University(香港理工大学健康科技与信息学系、数据科学与人工智能系)

专题命中 领域大模型 :foundation model(abstract);分类 cs.AI、cs.LG

AI总结 提出CHoE方法,通过结构条件专家网络和结构感知路由机制,实现跨域异构图提示学习,在少样本跨域任务中优于基线方法。

Comments accepted by IJCAI 2026, 9 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.16864 2026-06-08 cs.LG cs.AI math.DS 版本更新 62%

Position: A Dynamical Systems Perspective is Needed to Advance Time Series Modeling

立场:需要动力系统视角以推进时间序列建模

Daniel Durstewitz, Christoph Jürgen Hemmer, Florian Hess, Charlotte Ricarda Doll, Lukas Eisenmann

机构 * University of Tübingen(图宾根大学)

专题命中 领域大模型 :foundation model(abstract);分类 cs.AI、cs.LG

AI总结 本文主张时间序列建模需引入动力系统视角,通过重构底层DS实现更优预测,并讨论其理论优势与具体建议。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.06650 2026-06-08 cs.HC 新提交 50%

LinkNav: Surfacing Interconnected Information in Scientific Articles

LinkNav:在科学文章中呈现互联信息

Sebastian Joseph, Jennifer Healey, Junyi Jessy Li, Ani Nenkova

专题命中 领域大模型 :language model(abstract)

AI总结 提出LinkNav系统,通过语言模型生成阅读时的问题并搜索文档内答案,建立非相邻段落间的显式连接,提升学术论文阅读体验。

Comments 10 pages, 3 figures, ACL 2026 (Demo Track)

详情

展开后加载摘要…

URL PDF HTML 收藏