arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-03-06 至 2026-03-06 共收录 13 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 13 篇

2603.04647 2026-03-06 cs.CL 88%

Coordinated Semantic Alignment and Evidence Constraints for Retrieval-Augmented Generation with Large Language Models

协同语义对齐与证据约束在大型语言模型检索增强生成中的应用

Xin Chen, Saili Uday Gadgil, Jiarong Qiu

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 本文提出一种通过协同语义对齐与证据约束提升检索增强生成效果的方法,增强事实可靠性和生成流畅性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04480 2026-03-06 q-bio.QM cs.LG 88%

AbAffinity: A Large Language Model for Predicting Antibody Binding Affinity against SARS-CoV-2

AbAffinity:一种用于预测抗SARS-CoV-2病毒抗体结合亲和力的大语言模型

Faisal Bin Ashraf, Animesh Ray, Stefano Lonardi

机构 * Department of Computer Science and Engineering,University of California, Riverside(加州大学河滨分校计算机科学与工程系) Riggs School of Applied Life Sciences, Keck Graduate Institute, Claremont(克劳斯研究生院应用生命科学学院)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.LG

AI总结 AbAffinity是一种通过大语言模型预测抗SARS-CoV-2病毒抗体结合亲和力的新型方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05432 2026-03-06 cs.CL cs.AI cs.LG 85%

Ensembling Language Models with Sequential Monte Carlo

用序列蒙特卡洛方法融合语言模型

Robin Shing Moon Chan, Tianyu Liu, Samuel Kiegeland, Clemente Pasti, Jacob Hoover Vigly, Timothy J. O'Donnell, Ryan Cotterell, Tim Vieira

机构 * ETH Zürich(苏黎世联邦理工学院) McGill University(麦吉尔大学) Canada CIFAR AI Chair(加拿大 CIFAR 人工智能主席)

专题命中 其他LLM :language model(title,abstract);prompting(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出了一种融合语言模型的统一框架,通过字节级序列蒙特卡洛算法实现不同词汇表模型的融合采样,提升了结构化文本生成任务的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01870 2026-03-06 q-bio.NC cs.AI cs.LG 81%

CytoNet: A Foundation Model for the Human Cerebral Cortex at Cellular Resolution

CytoNet:人类大脑皮层细胞分辨率的基础模型

Christian Schiffer, Zeynep Boztoprak, Jan-Oliver Kropp, Julia Thönnißen, Katia Berr, Hannah Spitzer, Katrin Amunts, Timo Dickscheid

机构 * Institute of Neuroscience and Medicine (INM-1), Research Centre Jülich(神经科学与医学研究所(INM-1),贾格尔研究中心) Helmholtz AI, Research Centre Jülich(海德堡人工智能,贾格尔研究中心) Institute for Brain Research, University Hospital Düsseldorf(脑研究所在杜塞尔多夫大学医院) Institute of Computational Biology, Computational Health Center, Helmholtz Munich(计算生物学研究所,计算健康中心,海德堡慕尼黑) Institute for Stroke and Dementia Research (ISD), LMU University Hospital, LMU Munich(中风与痴呆研究所以(ISD),慕尼黑大学医院,慕尼黑大学) Computer Vision, Institute for Computational Visualistics, University of Koblenz(计算机视觉,计算视觉研究所,科布伦兹大学)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI、cs.LG

AI总结 CytoNet通过训练100万张显微图像片段,建立了人类大脑皮层细胞分辨率的基础模型,用于分析微架构并链接细胞结构与功能组织。

Comments 42 pages, 10 figures, 7 tables. Extended version with functional decoding

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04514 2026-03-06 cs.AI 79%

Progressive Refinement Regulation for Accelerating Diffusion Language Model Decoding

逐步细化调节以加速扩散语言模型解码

Lipeng Wan, Jianhui Gu, Junjie Ma, Jianguo Huang, Shiguang Sun, Siyuan Li, Xuguang Lan

机构 * Department of Artificial Intelligence, Xi'an Jiaotong University(人工智能系,西安交通大学) College of Computing and Data Science, Nanyang Technological University(计算与数据科学学院,南洋理工大学) School of Computer Science and Technology, Harbin Institute of Technology(计算机科学与技术学院,哈尔滨工业大学)

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

AI总结 逐步细化调节通过动态控制细化规则提升扩散语言模型解码效率并保持生成质量。

Comments 19 pages, 10 figures, Code available upon publication

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.04548 2026-03-06 cs.CL cs.AI 79%

Identifying Good and Bad Neurons for Task-Level Controllable LLMs

识别用于任务级可控大语言模型的好神经元和坏神经元

Wenjie Li, Guansong Pang, Hezhe Qiao, Debin Gao, David Lo

机构 * Singapore Management University(新加坡管理大学) ShanghaiTech University(上海科技大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 NeuronLLM通过对比学习识别任务级可控大语言模型中的好神经元和坏神经元,以提升任务表现和理解能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05240 2026-03-06 cs.AI 77%

GCAgent: Enhancing Group Chat Communication through Dialogue Agents System

GCAgent:通过对话代理系统增强群聊交流

Zijie Meng, Zheyong Xie, Zheyu Ye, Chonggang Lu, Zuozhu Liu, Zihan Niu, Yao Hu, Shaosheng Cao

机构 * Zhejiang University(浙江大学) Xiaohongshu Inc.(小红书公司) University of Science and Technology of China(中国科学技术大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 GCAgent通过集成娱乐和实用导向的对话代理系统,提升群聊交流效果,实现多参与者对话的高效管理与增强。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04735 2026-03-06 cs.AI cs.CL 73%

Solving an Open Problem in Theoretical Physics using AI-Assisted Discovery

用人工智能辅助发现解决理论物理中的一个开放问题

Michael P. Brenner, Vincent Cohen-Addad, David Woodruff

机构 * Google Research(谷歌研究) School of Engineering and Applied Sciences, Harvard University(工程与应用科学学院,哈佛大学) School of Computer Science, Carnegie Mellon University(计算机科学学院,卡内基梅隆大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本文提出了一种结合大型语言模型和树搜索框架的混合系统,通过AI辅助发现解决了理论物理中宇宙弦引力辐射功率谱的开放问题,并推导出新的解析解。

Comments 22 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25128 2026-03-06 astro-ph.IM astro-ph.HE 71%

Towards a foundation model for astrophysical source detection: An End-to-End Gamma-Ray Data Analysis Pipeline Using Deep Learning

迈向天体源检测的基础模型:一种使用深度学习的端到端伽马射线数据分析流水线

Judit Pérez-Romero, Saptashwa Bhattacharyya, Sascha Caron, Dmitry Malyshev, Rodney Nicolas, Giacomo Principe, Zoja Rokavec, Roberto Ruiz de Austri, Danijel Skočaj, Fiorenzo Stoppa, Domen Tabernik, Gabrijela Zaharijas

专题命中 其他LLM :foundation model(title)

AI总结 本文提出了一种基于深度学习的端到端伽马射线数据分析流水线,扩展了AutoSourceID方法以应用于CTAO模拟数据,为天体源检测提供灵活且鲁棒的解决方案。

Comments 6 pages, 3 figures, presented at EuCAIFCon 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.11432 2026-03-06 cs.CL 70%

The unreasonable effectiveness of pattern matching

模式匹配的不合理有效性

Gary Lupyan, Blaise Agüera y Arcas

机构 * Department of Psychology University of Wisconsin–Madison(心理学系 威斯康星大学麦迪逊分校) Paradigms of Intelligence Team Google(智能范式团队 谷歌)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 大型语言模型通过模式匹配有效理解随机替换实词的虚构文本,揭示了其在语言处理中的核心能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04613 2026-03-06 cs.HC 67%

Beyond Anthropomorphism: a Spectrum of Interface Metaphors for LLMs

超越拟人化:面向大语言模型的界面隐喻光谱

Jianna So, Connie Cheng, Sonia Krishna Murthy

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出将拟人化作为设计变量,通过反拟人化到超拟人化的隐喻光谱,引导界面设计从优化可用性转向鼓励批判性参与。

Comments Extended Abstracts of the 2026 CHI Conference on Human Factors in Computing Systems

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05609 2026-03-06 cs.CL cs.LG 62%

New Insights into Optimal Alignment of Acoustic and Linguistic Representations for Knowledge Transfer in ASR

对ASR中知识转移的声学与语言表示最优对齐的新见解

Xugang Lu, Peng Shen, Hisashi Kawai

机构 * National Institute of Information and Communications Technology, Japan(日本信息与通信技术研究所)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.LG

AI总结 本文提出了一种基于不平衡最优传输的对齐模型,用于改进ASR中声学与语言表示的对齐,提升知识转移效果。

Comments Accepted to ICASSP 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.15738 2026-03-06 cs.LG cs.AI cs.DC 54%

Parallel Split Learning with Global Sampling

并行分割学习与全局采样

Mohammad Kohankhaki, Ahmad Ayad, Mahdi Barhoush, Anke Schmeink

专题命中 其他LLM :分类 cs.AI、cs.LG;large language model(comments);language model(comments)

AI总结 GPSL通过全局采样解决并行分割学习中的批量大小和非IID数据问题,实现稳定优化和集中式准确性,同时减少训练时间。

Comments Accepted at the 2025 IEEE 3rd International Conference on Foundation and Large Language Models (FLLM). This version corresponds to the accepted manuscript

详情

展开后加载摘要…

URL PDF HTML 收藏