arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-05-05 至 2026-05-05 共收录 15 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 15 篇

2605.01097 2026-05-05 cs.CL cs.AI 87%

Interpretable Difficulty-Aware Knowledge Tracing in Tutor-Student Dialogues

可解释的难度感知对话知识追踪

Shuyan Huang, Alexander Scarlatos, Jaewook Lee, Andrew Lan

机构 * University of Massachusetts Amherst(马萨诸塞大学阿默斯特分校)

专题命中 其他LLM :LLM(summary_cn,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本文提出基于LLM的可解释难度感知对话知识追踪框架,通过建模学生能力与问题难度,提升对话场景下的个性化支持能力。

Comments 11 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.01167 2026-05-05 cs.LG cs.AI 82%

Minimizing Collateral Damage in Activation Steering

最小化激活引导中的附带损害

Tam Nguyen, Tu Anh Nguyen, Sina Alemohammad, Richard G. Baraniuk

机构 * Department of Electrical \& Computer Engineering, Rice University, Houston, USA Department of Computational Applied Mathematics, Rice University, Houston, USA Department of Electrical \& Computer Engineering, The University of Texas at Austin, Austin, USA

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出了一种基于约束优化的框架,通过数学形式化附带损害并优化激活变化,以更精确地控制大语言模型行为,同时减少对无关任务性能的影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00844 2026-05-05 cs.CY cs.AI 81%

The Oracle's Fingerprint: Correlated AI Forecasting Errors and the Limits of Bias Transmission

预言者指纹:相关AI预测误差与偏见传递的极限

Theodor Spiro

机构 * Theodor Spiro

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究揭示大型语言模型预测误差的高度相关性,以及人类群体预测受其影响的机制,发现AI系统与人类偏见模式的相似性及传递效应。

Comments 23 pages, 3 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00827 2026-05-05 cs.DC cs.AI cs.SE 81%

Separating Intelligence from Execution: A Workflow Engine for the Model Context Protocol

将智能与执行分离:一种用于模型上下文协议的工作流引擎

Abhinav Singh Parmar

机构 * Infosys

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出一种MCP原生编排层,将智能决策与执行分离,通过工作流蓝图减少token消耗,提升执行效率。

Comments 16 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00850 2026-05-05 physics.ao-ph cs.AI cs.LG eess.IV 81%

Earth System Foundation Model (ESFM): A unified framework for heterogeneous data integration and forecasting

地球系统基础模型(ESFM):一个用于异质数据整合和预测的统一框架

Firat Ozdemir, Yun Cheng, Salman Mohebi, Fanny Lehmann, Simon Adamov, Zhenyi Zhang, Leonardo Trentini, Dana Grund, Oliver Fuhrer, Torsten Hoefler, Siddhartha Mishra, Sebastian Schemm, Benedikt Soja, Mathieu Salzmann

机构 * Swiss Data Science Center (SDSC), ETH Zurich & EPFL(瑞士数据科学中心(SDSC)、苏黎世联邦理工学院与洛桑联邦理工学院) ETH AI Center, ETH Zurich(苏黎世联邦理工学院人工智能中心) Federal Office for Meteorology and Climatology MeteoSwiss(联邦气象与气候办公室MeteoSwiss) Institute of Geodesy and Photogrammetry, ETH Zurich(苏黎世联邦理工学院测绘与摄影测量研究所) Computational and Applied Mathematics Laboratory, ETH Zurich(苏黎世联邦理工学院计算与应用数学实验室) Institute for Atmospheric and Climate Science (IAC), ETH Zurich(大气与气候科学研究所(IAC)、苏黎世联邦理工学院) Scalable Parallel Computing Lab, ETH Zurich(可扩展并行计算实验室、苏黎世联邦理工学院) Dep. of Applied Mathematics and Theoretical Physics, University of Cambridge(应用数学与理论物理系、剑桥大学)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI、cs.LG

AI总结 本文提出ESFM,一个基于3D Swin UNet的开放模型,通过扩展编码方案和轴向注意力机制,实现对多源异质数据的整合与预测,提升气候科学应用性能。

Comments ESFM is available on https://github.com/swiss-ai/ESFM. 48 pages, 29 figures, 18 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00968 2026-05-05 eess.SP cs.AI 79%

Adaptive 3D-RoPE: Physics-Aligned Rotary Positional Encoding for Wireless Foundation Models

自适应3D-RoPE:用于无线基础模型的物理对齐旋转位置编码

Chenyu Zhang, Xinchen Lyu, Chenshan Ren, Shuhan Liu, Qimei Cui

机构 * National Engineering Research Center for Mobile Network Technologies, Beijing University of Posts and Telecommunications(中国移动网络技术国家工程研究中心,北京邮电大学) Department of Broadband Communication, Pengcheng Laboratory(宽带通信部,鹏城实验室) Key Laboratory of Ethnic Language Intelligent Analysis and Security Governance of MOE, Minzu University of China(教育部民族语言智能分析与安全治理重点实验室,中央民族大学) China Telecom Corporation Limited Gansu Branch(中国电信集团甘肃分公司)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI

AI总结 本文提出自适应3D-RoPE,通过整合可学习的3D频率银行和轻量级通道条件控制器,解决无线基础模型中位置编码的静态问题,提升尺度外推和零样本泛化性能。

Comments 13 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.21446 2026-05-05 cs.AI cs.CL cs.MA cs.SI 73%

AI-Gram: When Visual Agents Interact in a Social Network

AI-Gram:当视觉代理在社交网络中互动

Andrew Shin

机构 * Faculty of Science and Technology(科学与技术学部)

专题命中 其他LLM :LLM(abstract,abstract_cn);分类 cs.CL、cs.AI

AI总结 AI-Gram是一个自主运行的社交平台,研究AI在大规模社交动态中的行为,通过视觉代理生成和回应视觉内容,揭示三阶段动态:视觉回复链、审美主权和审美复调。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.01302 2026-05-05 cs.CL cs.IR 70%

Beyond Semantic Relevance: Counterfactual Risk Minimization for Robust Retrieval-Augmented Generation

超越语义相关性:为鲁棒检索增强生成的反事实风险最小化

Peiyang Liu, Qiang Yan, Ziqiang Cui, Di Liang, Xi Wang, Wei Ye

机构 * National Engineering Research Center for Software Engineering, Peking University(软件工程国家工程研究中心,北京大学) City University of Hong Kong(香港城市大学) Tencent Technology(腾讯科技) Peking University(北京大学)

专题命中 其他LLM :LLM(abstract,abstract_cn);分类 cs.CL

AI总结 本文提出CoRM-RAG框架,通过因果干预模拟用户偏差来提升检索鲁棒性,在对抗性场景中优于传统检索器和重排序器。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00963 2026-05-05 cs.RO cs.AI 70%

Ablation Study of Multimodal Perception, Language Grounding, and Control for Human-Robot Interaction in an Object Detection and Grasping Task

多模态感知、语言接地与控制在物体检测与抓取任务中的人机交互消融研究

Zi Tian, Guanting Shen

机构 * Dalian University of Technology(大连理工大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文通过消融研究探讨多模态感知、语言接地和控制模块对端到端性能的影响,评估最佳组合以优化执行时间和成功率。

Comments 10 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.23940 2026-05-05 cs.SE cs.AI 70%

Constraint-Guided Multi-Agent Decompilation for Executable Binary Recovery

基于约束的多智能体反编译用于可执行二进制恢复

Yifan Zhang, Xiaohan Wang, Yueke Zhang, Yu Huang, Kevin Leach

机构 * Vanderbilt University(范德比大学)

专题命中 其他LLM :LLM(abstract,abstract_cn);分类 cs.AI

AI总结 本文提出一种多智能体框架,通过多级约束引导反编译(MCGD)将反编译代码转换为可执行源代码,提升反编译结果的可执行性与正确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.01533 2026-05-05 cs.SE 67%

Genetic Programming for Self-Adaptive Auto-Scaling of Microservices

基于遗传编程的微服务自适应自动扩展

Jia Li, Mehrdad Sabetzadeh, Shiva Nejati

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出AutoSLO框架,利用遗传编程动态调整微服务副本以满足SLO并减少资源消耗,通过持续监控反馈循环实现主动预防SLO违规。

Comments This paper has been accepted into the 18th Symposium on Search-Based Software Engineering 2026 (SSBSE 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.01365 2026-05-05 cs.CV cs.RO 67%

VoxAfford: Multi-Scale Voxel-Token Fusion for Open-Vocabulary 3D Affordance Detection

VoxAfford:多尺度体素-标记融合用于开放词汇3D affordance检测

Haowen Sun, Shaolong Zhang, Mingyang Li, Chengzhong Ma, Xinzhe Chen, Qiongjie Cui, Xingyu Chen, Zeyang Liu, Xuguang Lan

机构 * National Key Laboratory of Human-Machine Hybrid Augmented Intelligence, Institute of Artificial Intelligence and Robotics, Xi’an Jiaotong University(人机混合增强智能国家级实验室,人工智能与机器人研究所,西安交通大学)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出VoxAfford,通过多尺度几何特征增强输出标记,提升3D affordance检测的定位精度,实验显示mIoU提升8%,并验证了零样本迁移能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.01611 2026-05-05 cs.CY cs.AI cs.LG 62%

The Case for ESM3 as a General-Purpose AI Model with Systemic Risk Under the EU AI Act

ESM3作为具有系统性风险的通用人工智能模型在欧盟人工智能法案中的案例

Taro Qureshi, Jacob Griffith, Koen Holtman, Marcel Mir Teijeiro, Ze Shen Chin, Rokas Gipiškis

机构 * AI Standards Lab(人工智能标准实验室) Vilnius University(维尔纽斯大学) Northeastern University London(伦敦东北大学) London School of Economics(伦敦政治经济学院)

专题命中 其他LLM :foundation model(abstract);分类 cs.AI、cs.LG

AI总结 本文探讨ESM3等前沿生物基础模型在欧盟人工智能法案下的监管问题,分析其是否受通用人工智能模型系统性风险义务约束,并提出改进建议。

Comments 8 pages, 1 figure, Technical AI Safety Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.01637 2026-05-05 cs.LG cs.CC cs.DM math.CO 61%

The Banach-Butterfly Invariant: Influence-Adaptive Walsh Geometry for Ternary Polynomial Threshold Functions

Banach-Butterfly 不变量:适应影响的Walsh几何用于三元多项式阈值函数

Gorgi Pavlov

机构 * Lehigh University(莱德大学) Johnson and Johnson(强生公司)

专题命中 其他LLM :LLM(abstract_cn,comments);分类 cs.LG

AI总结 本文提出Banach-Butterfly不变量,用于三元多项式阈值函数的适应影响的Walsh几何分析,通过影响向量的Schur凸性分离函数,并证明其作为收缩不变量的性质。

Comments 21 pages, 3 figures. Theory paper; LLM-application companion in preparation. Code, certificates, and 616,126 NPN-canonical n=5 representatives in supplementary repository

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00948 2026-05-05 q-bio.QM cs.AI 57%

Co-Generative De Novo Functional Protein Design

协同生成的从头功能性蛋白质设计

Xinrui Chen, Yizhen Luo, Siqi Fan, Zaiqing Nie

机构 * Institute for AI Industry Research (AIR), Tsinghua University(人工智能产业研究院(AIR),清华大学) Department of Computer Science(计算机科学系) Technology, Tsinghua University(技术,清华大学)

专题命中 其他LLM :language model(abstract);分类 cs.AI

AI总结 本文提出CodeFP模型,通过同时解码序列和结构令牌,提升功能性与折叠性的实现。模型利用功能局部结构增强语义编码,并引入辅助功能监督以缓解训练模糊性,实验显示在功能一致性和折叠性上均优于基线。

详情

展开后加载摘要…

URL PDF HTML 收藏