arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-04-13 至 2026-04-13 共收录 12 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 12 篇

2604.06737 2026-04-13 cs.CL cs.AI 90%

WisdomInterrogatory (LuWen): An Open-Source Legal Large Language Model Technical Report

WisdomInterrogatory (LuWen): 一种开源法律大语言模型技术报告

Yiquan Wu, Yuhang Liu, Yifei Liu, Ang Li, Siying Zhou, Kun Kuang, Fei Wu

机构 * Zhejiang University(浙江大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);foundation model(abstract);分类 cs.CL、cs.AI

AI总结 本文提出基于Baichuan模型的开源中文法律语言模型LuWen,通过持续预训练、监督微调和检索增强生成技术,提升法律领域任务表现。

Comments 10 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.08555 2026-04-13 cs.CL 89%

SynDocDis: A Metadata-Driven Framework for Generating Synthetic Physician Discussions Using Large Language Models

SynDocDis:一种基于元数据的生成合成医生讨论的框架

Beny Rubinstein, Sergio Matos

机构 * University of Aveiro(阿威罗大学) IEETA, DETI, LASI, University of Aveiro(阿威罗大学 IEETA、DETI、LASI)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);prompting(abstract);分类 cs.CL

AI总结 SynDocDis通过结合结构化提示技术和隐私保护的去标识病例元数据,生成临床准确的医生间对话,经九个肿瘤学和肝病学场景评估,显示优异的沟通效果和医学内容质量。

Journal ref In: Valente de Oliveira, J., Leite, J., Rodrigues, J., Dias, J., Cardoso, P. (eds) Progress in Artificial Intelligence. EPIA 2025. Lecture Notes in Computer Science(), vol 16121. Springer, Cham

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.08566 2026-04-13 cs.CL cs.LG 88%

Sentiment Classification of Gaza War Headlines: A Comparative Analysis of Large Language Models and Arabic Fine-Tuned BERT Models

加沙战争头条的情感分类:大型语言模型与阿拉伯语微调BERT模型的比较分析

Amr Eleraqi, Hager H. Mustafa, Abdul Hadi N. Ahmed

机构 * Anmat Media

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.LG

AI总结 本文通过加沙战争头条数据,比较大型语言模型与阿拉伯语微调BERT模型在情感分类中的表现,揭示模型架构对情感解读的影响及系统性差异。

Comments 45 pages, 6 figures (including diagrams), 8 tables. Dataset available at this https URL . Previously posted at https://dataverse.harvard.edu/dataset.xhtml?persistentId=doi:10.7910/DVN/FFENX3

Journal ref SSRN (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.19275 2026-04-13 cs.CL cs.AI 88%

Improving Automatic Summarization of Radiology Reports through Mid-Training of Large Language Models

通过在大语言模型中段训练改进放射报告的自动摘要

Mengxian Lyu, Cheng Peng, Ziyi Chen, Mengyuan Zhang, Jieting Li Lu, Yonghui Wu

机构 * University of Florida(佛罗里达大学) Department of Health Outcomes and Biomedical Informatics(健康结果与生物医学信息学系) Department of Engineering Education(工程教育系)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文提出在大语言模型中段训练的方法,通过三种策略提升放射报告摘要效果,实验表明GatorTronT5-Radio在文本和事实性指标上表现最佳,且在少样本学习中表现更优。

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.08691 2026-04-13 cs.SI cs.AI 85%

AgentSociety: Large-Scale Simulation of LLM-Driven Generative Agents Advances Understanding of Human Behaviors and Society

AgentSociety: 基于大语言模型的生成代理大规模模拟推动对人类行为与社会的理解

Jinghua Piao, Yuwei Yan, Jun Zhang, Nian Li, Junbo Yan, Xiaochong Lan, Zhihong Lu, Zhiheng Zheng, Jing Yi Wang, Di Zhou, Chen Gao, Fengli Xu, Fang Zhang, Ke Rong, Jun Su, Yong Li

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出AgentSociety平台,通过大规模模拟人类行为和社会动态,探索社会问题如极化、虚假信息传播等,验证其在社会科学研究中的应用价值。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.08956 2026-04-13 cs.CV cs.LG 85%

Low-Data Supervised Adaptation Outperforms Prompting for Cloud Segmentation Under Domain Shift

低数据监督适应在云分割领域优于提示

Harshith Kethavath, Weiming Hu

机构 * University of Georgia(佐治亚大学)

专题命中 领域大模型 :prompting(title,abstract);language model(abstract);pretraining(abstract);分类 cs.LG

AI总结 本文研究了在遥感影像云分割任务中,低数据监督微调优于提示方法,发现使用少量标注数据可显著提升性能,而提示方法效果有限。

Comments 10 pages, 6 figures, to be published in EarthVision @ CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.09067 2026-04-13 cs.CV cs.AI 79%

Enhancing the Safety of Medical Vision-Language Models by Synthetic Demonstrations

通过合成演示增强医疗视觉-语言模型的安全性

Zhiyu Xue, Reza Abbasi-Asl, Ramtin Pedarsani

机构 * UC Santa Barbara(加州大学圣塔芭芭拉分校) UC San Francisco(加州大学旧金山分校)

专题命中 领域大模型 :language model(title,abstract);分类 cs.AI

AI总结 本文提出一种新的推理时防御策略,通过合成临床演示提升模型安全性,同时平衡安全与性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.08815 2026-04-13 cs.CV 78%

Towards Responsible Multimodal Medical Reasoning via Context-Aligned Vision-Language Models

通过上下文对齐的视觉-语言模型实现负责任的多模态医疗推理

Sumra Khan, Sagar Chhabriya, Aizan Zafar, Sheeraz Arif, Amgad Muneer, Anas Zafar, Shaina Raza, Rizwan Qureshi

机构 * Salim Habib University(萨利姆·哈比卜大学) Institute of Business Administration Sukkur(苏库尔工商管理学院) University of Central Florida(中佛罗里达大学) The University of Texas MD Anderson Cancer Center(德克萨斯大学MD安德森癌症中心) Toronto Metropolitan University(多伦多都会大学) Vector Institute(向量研究所)

专题命中 领域大模型 :language model(title,abstract)

AI总结 本文提出一种上下文对齐的框架,通过整合多种临床证据提升医疗多模态推理的可靠性与可信度,实验显示其在胸部X光数据集上提升了判别性能并减少幻觉关键词。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.09037 2026-04-13 cs.CV cs.CL cs.HC 70%

SiMing-Bench: Evaluating Procedural Correctness from Continuous Interactions in Clinical Skill Videos

SiMing-Bench:从连续交互评估临床技能视频中的过程正确性

Xiyang Huang, Jiawei Lin, Keying Wu, Jiaxin Huang, Kailai Yang, Renxiong Wei, Cheng zeng, Jiayi Xiang, Ziyan Kuang, Min Peng, Qianqian Xie, Sophia Ananiadou

机构 * School of Artificial Intelligence, Wuhan University(武汉大学人工智能学院) Center for Language and Information Research, Wuhan University(武汉大学语言与信息研究中心) Southwest Jiaotong University(西南交通大学) MBZUAI(穆罕默德·本·扎耶德人工智能大学) The University of Manchester(曼彻斯特大学) Zhongnan Hospital of Wuhan University(武汉大学中南医院)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 SiMing-Bench旨在评估多模态大语言模型在临床技能视频中通过连续交互更新过程状态的能力,通过医师标注的数据集和标准化评分标准,揭示模型在过程级判断上的不足。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.20638 2026-04-13 cs.SD cs.CV cs.MM eess.AS 67%

Music Audio-Visual Question Answering Requires Specialized Multimodal Designs

音乐音频视觉问答需要专门的多模态设计

Wenhao You, Xingjian Diao, Wenjun Huang, Chunhui Zhang, Keyi Kong, Weiyi Wu, Chiyu Ma, Zhongyu Ouyang, Tingxuan Wu, Ming Cheng, Soroush Vosoughi, Jiang Gui

机构 * University of Waterloo(滑铁卢大学) Dartmouth College(达特茅斯学院) UC Irvine(加州大学尔湾分校) New York University(纽约大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract)

AI总结 本文探讨了音乐音频视觉问答任务中多模态设计的必要性,指出需专门的输入处理、空间时间架构和音乐特定建模策略以应对连续密集的音频视觉内容和复杂时间动态。

Comments Accepted to Annual Meeting of the Association for Computational Linguistics (ACL 2026). The first two authors contributed equally

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.09501 2026-04-13 cs.CL 61%

You Can't Fight in Here! This is BBS!

你不能在这里战斗!这是BBS!

Richard Futrell, Kyle Mahowald

机构 * University of California Irvine(加州大学尔湾分校) The University of Texas at Austin(德克萨斯大学奥斯汀分校)

专题命中 领域大模型 :language model(abstract,comments);分类 cs.CL

AI总结 本文通过讨论语言模型在语言科学中的作用,揭示了字符串统计谬误和'好到不能再好'假设等核心问题,呼吁建立更广阔的语言科学研究计划。

Comments Accepted at Behavioral and Brain Sciences as a response to the commentaries to the accepted target article "How Linguistics Learned to Stop Worrying and Love the Language Models", whose preprint appears here: arXiv:2501.17047

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.05529 2026-04-13 cs.AI 57%

ActivityEditor: Learning to Synthesize Physically Valid Human Mobility

ActivityEditor: 学习合成物理有效的行人移动

Chenjie Yang, Yutian Jiang, Anqi Liang, Wei Qi, Chenyu Wu, Junbo Zhang

机构 * Southwest Jiaotong University(西南交通大学) HKUST (Guangzhou)(香港科技大学(广州)) Shanghai Jiao Tong University(上海交通大学) Tsinghua University(清华大学) JD Technology(京东科技)

专题命中 领域大模型 :LLM(abstract);分类 cs.AI

AI总结 本文提出ActivityEditor,一种双LLM代理框架,用于零样本跨区域轨迹生成。通过意图生成代理和编辑代理协同工作,结合强化学习和物理约束,实现高保真轨迹生成。

详情

展开后加载摘要…

URL PDF HTML 收藏