arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-03-06 至 2026-03-06 共收录 24 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 24 篇

2503.15664 2026-03-06 cs.CL 89%

Enhancing Pancreatic Cancer Staging with Large Language Models: The Role of Retrieval-Augmented Generation

利用大型语言模型增强胰腺癌分期:检索增强生成的作用

Hisashi Johno, Yuki Johno, Akitomo Amakawa, Junichi Sato, Ryota Tozuka, Atsushi Komaba, Hiroaki Watanabe, Hiroki Watanabe, Chihiro Goto, Hiroyuki Morisaka, Hiroshi Onishi, Kazunori Nakamoto

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL

AI总结 利用RAG技术提升胰腺癌分期准确性,展示NotebookLM在临床诊断中的应用价值

Comments 11 pages, 6 figures, 2 tables, 6 supplementary files

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05278 2026-03-06 cs.SE 89%

A framework for assessing the capabilities of code generation of constraint domain-specific languages with large language models

一种评估大语言模型生成约束领域特定语言代码能力的框架

David Delgado, Lola Burgueño, Robert Clarisó

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract)

AI总结 本文提出一种评估框架,用于评估大语言模型生成约束领域特定语言代码的能力,发现LLMs在Python上的表现优于OCL和Alloy,并探讨了改进代码生成的方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19948 2026-03-06 cs.CL cs.AI cs.CY cs.HC cs.MA 88%

Assessing Risks of Large Language Models in Mental Health Support: A Framework for Automated Clinical AI Red Teaming

评估大型语言模型在心理健康支持中的风险:一种用于自动化临床AI红队测试的框架

Ian Steenstra, Paola Pedrelli, Weiyan Shi, Stacy Marsella, Timothy W. Bickmore

机构 * Northeastern University(东北大学) Harvard Medical School(哈佛医学院)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文提出了一种评估AI心理治疗师在心理健康支持中安全风险的框架,通过模拟测试发现AI在治疗中的潜在风险,并验证了交互式可视化工具的有效性。

Comments This paper is a condensed version of the first author's Ph.D. dissertation submitted to Northeastern University

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.10117 2026-03-06 cs.LG cs.CL 88%

Learning Virtual Machine Scheduling in Cloud Computing through Language Agents

通过语言代理学习云计算中的虚拟机调度

JieHao Wu, Ziwei Wang, Junjie Sheng, Wenhao Li, Xiangfeng Wang, Jun Luo

机构 * School of Computer Science and Technology, East China Normal University(东华大学计算机科学与技术学院) Antai College of Economics and Management, Shanghai Jiao Tong University(上海交通大学安泰经济管理学院) School of Computer Science and Technology, Tongji University(同济大学计算机科学与技术学院) Key Laboratory of Mathematics and Engineering Applications, MoE, East China Normal University(教育部数学与工程应用重点实验室)

专题命中 领域大模型 :language agent(title,abstract);LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文提出MiCo框架,通过语言代理解决云计算中的虚拟机调度问题,实现高效动态调度并验证其在大规模场景中的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04476 2026-03-06 cs.SE cs.PL 88%

iScript: A Domain-Adapted Large Language Model and Benchmark for Physical Design Tcl Script Generation

iScript: 一种领域适应的大型语言模型和基准,用于物理设计 Tcl 脚本生成

Ning Xu, Zhaoyang Zhang, Senlin Shu, Lei Qi, Jiaqi Lv, Wensuo Wang, Tianhao Zhao, Chao Zhang, Zhaoliang Yang, Xiangyu Li, Zhaorui Su, Jingshan Li, Xin Geng

专题命中 领域大模型 :large language model(title);language model(title);LLM(abstract);pretraining(abstract)

AI总结 iScript 是一种专门用于物理设计 Tcl 脚本生成的领域适应大语言模型,通过多阶段数据合成和两阶段训练策略提升脚本生成性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04743 2026-03-06 cs.IR cs.AI cs.CL 86%

DARE: Aligning LLM Agents with the R Statistical Ecosystem via Distribution-Aware Retrieval

DARE: 通过分布感知检索对齐LLM代理与R统计生态系统

Maojun Sun, Yue Wu, Yifei Xie, Ruijian Han, Binyan Jiang, Defeng Sun, Yancheng Yuan, Jian Huang

机构 * Department of Data Science and Artificial Intelligence, The Hong Kong Polytechnic University, Hong Kong SAR, China(数据科学与人工智能系,香港理工大学,香港特别行政区,中国) Department of Applied Mathematics, The Hong Kong Polytechnic University, Hong Kong SAR, China(应用数学系,香港理工大学,香港特别行政区,中国)

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 DARE通过整合数据分布信息提升R包检索效果,构建了面向R的LLM代理以实现更高效的统计分析任务。

Comments 24 pages,7 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.11791 2026-03-06 cs.IR 82%

LEXA: Legal Case Retrieval via Graph Contrastive Learning with Contextualised LLM Embeddings

通过具有上下文化LLM嵌入的图对比学习进行法律案例检索:LEXA

Yanran Tang, Ruihong Qiu, Yilun Liu, Xue Li, Zi Huang

专题命中 领域大模型 :LLM(title);large language model(abstract);language model(abstract)

AI总结 LEXA通过结合图对比学习和上下文化LLM嵌入,改进法律案例检索的结构信息利用与模型性能。

Comments arXiv admin note: substantial text overlap with arXiv:2312.11229

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27173 2026-03-06 cs.CE cs.AI cs.LG math.DS 81%

FMint-SDE: A Multimodal Foundation Model for Accelerating Numerical Simulation of SDEs via Error Correction

FMint-SDE:一种多模态基础模型,通过误差校正加速微分方程的数值模拟

Jiaxin Yuan, Haizhao Yang, Maria Cameron

机构 * University of Maryland(马里兰大学)

专题命中 领域大模型 :foundation model(title,abstract);分类 cs.AI、cs.LG

AI总结 FMint-SDE通过多模态学习实现误差校正,提升微分方程数值模拟的精度与效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04692 2026-03-06 cs.LG 79%

Engineering Regression Without Real-Data Training: Domain Adaptation for Tabular Foundation Models Using Multi-Dataset Embeddings

工程回归无需真实数据训练:使用多数据集嵌入进行表格基础模型的领域适应

Lyle Regenwetter, Rosen Yu, Cyril Picard, Faez Ahmed

机构 * Massachusetts Institute of Technology(麻省理工学院)

专题命中 领域大模型 :foundation model(title,abstract);分类 cs.LG

AI总结 本文提出一种无需真实数据训练的工程回归方法,通过多数据集嵌入技术提升基础模型在数据稀少领域的表现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07081 2026-03-06 cs.AI 79%

ClinNoteAgents: An LLM Multi-Agent System for Predicting and Interpreting Heart Failure 30-Day Readmission from Clinical Notes

ClinNoteAgents: 基于大语言模型的多智能体系统用于从临床笔记预测和解释心力衰竭30天再住院

Rongjia Zhou, Chengzhuo Li, Carl Yang, Jiaying Lu

机构 * Emory University(埃默里大学)

专题命中 领域大模型 :LLM(title,abstract);分类 cs.AI

AI总结 ClinNoteAgents通过多智能体系统从临床笔记中预测和解释心力衰竭30天再住院风险,提升医疗数据利用效率。

Comments 10 pages, 2 figures. Accepted to AMIA 2026 Informatics Summit (Student Paper Track)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.14882 2026-03-06 cs.CL 79%

Llama-Mimi: Exploring the Limits of Flattened Speech Language Modeling

Llama-Mimi:探索扁平化语音语言模型的极限

Issa Sugiura, Shuhei Kurita, Yusuke Oda, Ryuichiro Higashinaka

机构 * Kyoto University(京都大学) NII LLMC(日本信息处理学会大语言模型中心) National Institute of Informatics(国家信息研究所) Nagoya University(名古屋大学)

专题命中 领域大模型 :language model(title,abstract);分类 cs.CL

AI总结 Llama-Mimi通过将多级RVQ标记扁平化并使用Transformer解码器进行自回归建模,实现了在语音任务中的优越性能。

Comments 6 pages, 1 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05012 2026-03-06 cs.CV 78%

Tell2Adapt: A Unified Framework for Source Free Unsupervised Domain Adaptation via Vision Foundation Model

Tell2Adapt:通过视觉基础模型实现源无关无监督领域自适应的统一框架

Yulong Shi, Shijie Li, Ziyi Li, Lin Qi

专题命中 领域大模型 :foundation model(title,abstract)

AI总结 Tell2Adapt通过视觉基础模型实现源无关无监督领域自适应,利用上下文感知提示正则化和视觉合理性细化提升医学图像分割性能。

Comments Accepted by IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20643 2026-03-06 cs.CR 78%

CyberSleuth: Autonomous Blue-Team LLM Agent for Web Attack Forensics

CyberSleuth:自主蓝队LLM代理用于网络攻击取证

Stefano Fumero, Kai Huang, Matteo Boffa, Danilo Giordano, Marco Mellia, Dario Rossi

专题命中 领域大模型 :LLM(title,abstract)

AI总结 CyberSleuth通过LLM代理实现网络攻击自动化取证,展示了多代理专业化和有效设计在提升取证效率中的作用。

Comments Updated version - Added study on Malware Traffic Analysis

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02338 2026-03-06 cs.SE cs.RO 78%

Vision Language Model-based Testing of Industrial Autonomous Mobile Robots

基于视觉语言模型的工业自主移动机器人测试

Jiahui Wu, Chengjie Lu, Aitor Arrieta, Shaukat Ali, Thomas Peyrucain

机构 * Simula Research Laboratory and University of Oslo(Simula研究实验室和奥斯陆大学) Mondragon University(蒙dragon大学) Simula Research Laboratory(Simula研究实验室) PAL Robotics(PAL机器人技术)

专题命中 领域大模型 :language model(title,abstract)

AI总结 本文提出基于视觉语言模型的测试方法,用于生成违反功能和安全要求的机器人交互场景,以提高自主移动机器人在复杂环境中的安全性和可靠性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01919 2026-03-06 cs.CR cs.AI cs.SE 77%

Real Money, Fake Models: Deceptive Model Claims in Shadow APIs

真实的钱,假的模型:影子API中的欺骗性模型声明

Yage Zhang, Yukun Jiang, Zeyuan Chen, Michael Backes, Xinyue Shen, Yang Zhang

机构 * CISPA Helmholtz Center for Information Security(CISPA海德堡信息安全研究中心)

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文首次系统审计官方LLM API与影子API,揭示了影子API在性能、安全性和身份验证方面的欺骗行为,影响科研可重复性和有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18929 2026-03-06 cs.IR cs.AI 77%

Give Users the Wheel: Towards Promptable Recommendation Paradigm

让用户掌控方向:迈向可提示的推荐范式

Fuyuan Lyu, Chenglin Luo, Qiyuan Zhang, Yupeng Hou, Haolun Wu, Xing Tang, Xue Liu, Jin L. C. Guo, Xiuqiang He

机构 * McGill \& Mila - Quebec AI Institute Montreal Canada Shenzhen Technological University Shenzhen China University of California San Diego San Diego US McGill University Montreal Canada McGill \& Mila - Quebec AI Institute Shenzhen Technological University University of California San Diego McGill University

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出DPR模型,通过解耦的可提示顺序推荐框架,使传统推荐模型能够动态利用自然语言提示优化检索过程,提升推荐效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04741 2026-03-06 cs.AI cs.DB cs.IR cs.LG 73%

CONE: Embeddings for Complex Numerical Data Preserving Unit and Variable Semantics

CONE:用于复杂数值数据的嵌入方法,保留单位和变量语义

Gyanendra Shrestha, Anna Pyayt, Michael Gubanov

机构 * Florida State University(佛罗里达州立大学) University of South Florida(佛罗里达州立大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 CONE通过复合嵌入方法提升复杂数值数据的语义编码,实现跨领域高精度数值推理。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05129 2026-03-06 cs.AI cs.MA 70%

MedCoRAG: Interpretable Hepatology Diagnosis via Hybrid Evidence Retrieval and Multispecialty Consensus

MedCoRAG:通过混合证据检索和多专科共识进行可解释的肝病诊断

Zheng Li, Jiayi Xu, Zhikai Hu, Hechang Chen, Lele Cong, Yunyun Wang, Shuchao Pang

机构 * School of Cyber Science and Engineering, Nanjing University of Science and Technology(南京理工大学信息科学与工程学院) School of Artificial Intelligence, Jilin University(吉林大学人工智能学院) Department of Neurology, China-Japan Union Hospital of Jilin University(吉林大学中日联谊医院神经内科) Department of Anesthesiology, China-Japan Union Hospital of Jilin University(吉林大学中日联谊医院麻醉科) School of Computing, Macquarie University(麦考瑞大学计算机学院)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 MedCoRAG通过混合证据检索和多专科共识,提升肝病诊断的可解释性和准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12112 2026-03-06 cs.CR 67%

BRIDG-ICS: AI-Grounded Knowledge Graphs for Intelligent Threat Analytics in Industry~5.0 Cyber-Physical Systems

BRIDG-ICS:面向工业5.0 CPS的AI驱动知识图谱用于智能威胁分析

Padmeswari Nandiya, Ahmad Mohsin, Ahmed Ibrahim, Iqbal H. Sarker, Helge Janicke

专题命中 领域大模型 :large language model(abstract);language model(abstract)

AI总结 BRIDG-ICS通过AI驱动的知识图谱实现工业5.0中网络物理系统的智能威胁分析与韧性评估

Comments 44 Pages, To be published in Springer Cybersecurity Journal

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04763 2026-03-06 cs.CV cs.AI cs.LG 62%

Evaluating GPT-5 as a Multimodal Clinical Reasoner: A Landscape Commentary

评估GPT-5作为多模态临床推理者的有效性:领域评论

Alexandru Florea, Shansong Wang, Mingzhe Hu, Qiang Li, Zach Eidex, Luke del Balzo, Mojtaba Safari, Xiaofeng Yang

机构 * Department of Radiation Oncology, Winship Cancer Institute, Emory University School of Medicine(放射肿瘤科,Winship癌症研究所,埃默里大学医学院) Department of Biomedical Engineering, Georgia Institute of Technology(生物医学工程系,佐治亚理工学院)

专题命中 领域大模型 :foundation model(abstract);分类 cs.AI、cs.LG

AI总结 本文评估GPT-5在多模态临床推理中的表现,发现其在文本推理和部分视觉问答任务中优于GPT-4o,但在神经放射学和乳腺摄影等专业领域仍显不足。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05392 2026-03-06 cs.AI 57%

Legal interpretation and AI: from expert systems to argumentation and LLMs

法律解释与人工智能:从专家系统到论证与大语言模型

Václav Janeček, Giovanni Sartor

专题命中 领域大模型 :language model(abstract);分类 cs.AI

AI总结 本文探讨了人工智能在法律解释中的应用,从专家系统到论证和大语言模型,旨在提升法律解释的精确性和自动化水平。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05354 2026-03-06 cs.CL eess.AS 57%

Exploring the potential and limitations of Model Merging for Multi-Domain Adaptation in ASR

探索模型合并在ASR多领域适应中的潜力与限制

Carlos Carvalho, Francisco Teixeira, Thomas Rolland, Alberto Abad

机构 * INESC-ID & 2 Instituto Superior Técnico, Universidade de Lisboa, Portugal(1 INESC-ID 与 2 里斯本大学理工学院, 里斯本大学, 葡萄牙)

专题命中 领域大模型 :foundation model(abstract);分类 cs.CL

AI总结 本文提出BoostedTSV-M算法,通过奇异值提升缓解排名崩溃并提高数值稳定性,在多领域ASR中实现优于完整微调的性能,同时保持分布外泛化能力。

Comments submitted for review for INTERSPEECH2026 conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21739 2026-03-06 cs.SD cs.LG eess.AS 57%

Noise-to-Notes: Diffusion-based Generation and Refinement for Automatic Drum Transcription

噪声到音符:基于扩散的生成与细化用于自动鼓件转录

Michael Yeung, Keisuke Toyama, Toya Teramoto, Shusuke Takahashi, Tamaki Kojima

机构 * Sony Group Corporation(索尼集团)

专题命中 领域大模型 :foundation model(abstract);分类 cs.LG

AI总结 本研究提出N2N框架,利用扩散模型将音频条件高斯噪声转化为鼓件事件,结合退火伪Huber损失和音乐基础模型特征提升鲁棒性,实现自动鼓件转录的生成与细化。

Comments Accepted to ICASSP 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.07409 2026-03-06 cs.CL 57%

Computational Fact-Checking of Online Discourse: Scoring scientific accuracy in climate change related news articles

在线讨论的计算事实核查:在气候变化相关新闻文章中评估科学准确性

Tim Wittenborg, Constantin Sebastian Tremel, Markus Stocker, Sören Auer

机构 * L3S Research Center, Leibniz University Hanover(莱比锡大学汉诺威分校L3S研究中心) TIB - Leibniz Information Centre for Science and Technology(莱比锡科学与技术信息中心)

专题命中 领域大模型 :LLM(abstract);分类 cs.CL

AI总结 本文提出了一种半自动化的在线讨论事实核查方法,通过知识图谱分析评估气候变化相关新闻文章的科学准确性,并指出需要进一步完善事实知识图谱以支持科学公民讨论。

Comments 8 pages, 7 figures, accepted at ICKG 2025

Journal ref 2025 IEEE International Conference on Knowledge Graph (ICKG), Limassol, Cyprus, 2025, pp. 371-378

详情

展开后加载摘要…

URL PDF HTML 收藏