arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-03-11 至 2026-03-11 共收录 17 信号源:cs.CL, cs.AI, cs.LG

1. 指令微调 17 篇

2512.15943 2026-03-11 cs.AI 92%

Small Language Models for Efficient Agentic Tool Calling: Outperforming Large Models with Targeted Fine-tuning

小型语言模型用于高效的代理工具调用:通过针对性微调超越大模型

Polaris Jhandi, Owais Kazi, Shreyas Subramanian, Neel Sendas

专题命中 指令微调 :language model(title,abstract);small language model(title,abstract);LLM(abstract);large language model(abstract)

AI总结 本文通过针对性微调小型语言模型,在工具调用任务中超越大模型,展示了SLMs在成本优化和效率提升方面的潜力。

Comments Accepted at AAAI 2026 Workshop on Agentic AI Benchmarks and Applications for Enterprise Tasks

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.07503 2026-03-11 cs.SE 91%

Evaluating Large Language Models for Multilingual Vulnerability Detection at Dual Granularities

评估大型语言模型在双粒度多语言漏洞检测中的性能

Honglin Shu, Michael Fu, Junji Yu, Dong Wang, Chakkrit Tantithamthavorn, Junjie Chen, Yasutaka Kamei

专题命中 指令微调 :large language model(title,abstract);language model(title,abstract);LLM(abstract);instruction tuning(abstract)

AI总结 本研究评估了大型语言模型在多语言漏洞检测中的性能,发现GPT-4o在函数级和行级检测中表现优异,优于其他模型。

Comments 53 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09865 2026-03-11 cs.LG 88%

GAST: Gradient-aligned Sparse Tuning of Large Language Models with Data-layer Selection

GAST:基于梯度对齐的大型语言模型稀疏调优与数据层选择

Kai Yao, Zhenghan Song, Kaixin Wu, Mingjie Zhong, Danzhao Cheng, Zhaorui Tan, Yixin Ji, Penglei Gao

机构 * Ant Group(蚂蚁集团) Cornell University(康奈尔大学) University of Liverpool(利物浦大学) Soochow University(苏州大学) Cleveland Clinic Lerner Research Institution(克利夫兰诊所勒纳研究机构)

专题命中 指令微调 :large language model(title,abstract);language model(title,abstract);分类 cs.LG

AI总结 GAST通过同时进行数据和层维度的选择性微调,解决现有方法忽略数据对不同层贡献差异的问题,提升大型语言模型的微调效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.08754 2026-03-11 cs.LG cs.AI 86%

Hindsight Credit Assignment for Long-Horizon LLM Agents

长远时间LLM智能体的回顾信用分配

Hui-Ze Tan, Xiao-Wen Yang, Hao Chen, Jie-Jing Shao, Yi Wen, Yuteng Shen, Weihong Luo, Xiku Du, Lan-Zhe Guo, Yu-Feng Li

专题命中 指令微调 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 HCAPO通过整合回顾信用分配提升LLM智能体在长期任务中的探索效率和决策简洁性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09892 2026-03-11 cs.LG cs.AI cs.CL 85%

MSSR: Memory-Aware Adaptive Replay for Continual LLM Fine-Tuning

MSSR: 为持续LLM微调的内存感知自适应回放

Yiyang Lu, Yu He, Jianlong Chen, Hongyuan Zha

专题命中 指令微调 :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 MSSR通过估计样本级别的记忆强度并自适应地安排复习,有效缓解持续LLM微调中的灾难性遗忘问题,同时保持快速适应能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09154 2026-03-11 cs.CL 83%

Bioalignment: Measuring and Improving LLM Disposition Toward Biological Systems for AI Safety

Bioalignment: 评估和改进LLM对生物系统的倾向以提高AI安全性

Trent R Northen, Mingxun Wang

机构 * Bioaligned Labs(Bioaligned实验室) Lawrence Berkeley National Lab(伯克利国家实验室) Computer Science & Engineering Dept, UC Riverside(加州大学河滨分校计算机科学与工程系)

专题命中 指令微调 :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本研究通过微调改进LLM对生物解决方案的偏好,表明少量微调可提升模型对生物与合成方法的权衡能力。

Comments 17 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06656 2026-03-11 cs.CV cs.AI 83%

GameVerse: Can Vision-Language Models Learn from Video-based Reflection?

GameVerse: 视觉-语言模型能否通过视频反思学习?

Kuan Zhang, Dongchen Liu, Qiyue Zhao, Jinkun Hou, Xinran Zhang, Qinlei Xie, Miao Liu, Yiming Li

专题命中 指令微调 :language model(title,abstract);SFT(abstract);分类 cs.AI

AI总结 GameVerse通过反思-重试范式评估视觉-语言模型在视频反思中的学习能力,结合失败轨迹和专家教程提升模型性能。

Comments https://gameverse-bench.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.08954 2026-03-11 cs.AI cs.CL cs.DC cs.IR cs.LG 82%

A Consensus-Driven Multi-LLM Pipeline for Missing-Person Investigations

基于共识的多语言模型流水线用于失踪人员调查

Joshua Castillo, Ravi Mukkamala

专题命中 指令微调 :LLM(title,abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出了一种基于共识的多语言模型流水线,用于失踪人员调查,通过协调多个任务专用的 LLM 模型和共识引擎,提升信息提取和处理的效率与准确性。

Comments Accepted to CAC: Applied Computing & Automation Conferences 2026. 16 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05494 2026-03-11 cs.LG cs.AI cs.CL 80%

Censored LLMs as a Natural Testbed for Secret Knowledge Elicitation

被屏蔽的LLM作为秘密知识提取的自然测试场

Helena Casademunt, Bartosz Cywiński, Khoi Tran, Arya Jakkli, Samuel Marks, Neel Nanda

机构 * Harvard University(哈佛大学) Warsaw University of Technology(华沙理工大学) IDEAS Research Institute(IDEAS研究所) CentraleSupélec(中央理工学院) Anthropic(Anthropic公司)

专题命中 指令微调 :large language model(abstract);language model(abstract);prompting(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 研究利用被屏蔽的LLM测试秘密知识提取技术,发现开放式权重模型在诚实提取和谎言检测中存在局限性,且无法完全消除虚假回应。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.03477 2026-03-11 cs.CV cs.LG 79%

Fairness-Aware Fine-Tuning of Vision-Language Models for Medical Glaucoma Diagnosis

面向医疗青光眼诊断的公平性感知视觉语言模型微调

Zijian Gu, Yuxi Liu, Zhenhao Zhang, Song Wang

机构 * Department of Computer Science, University of Rochester, NY, USA(罗切斯特大学计算机科学系) Biostatistics and Health Data Science, School of Medicine, Indiana University, Indianapolis, IN, USA(印第安纳大学医学院生物统计学与健康数据科学系) Department of Computer Science, University of Central Florida, FL, USA(佛罗里达州立大学计算机科学系)

专题命中 指令微调 :language model(title,abstract);分类 cs.LG

AI总结 本文提出公平性感知的低秩适应方法,通过三种机制减少医疗影像诊断中的种族差异,实现高效公平的青光眼诊断。

Comments AMIA 2026 Amplify Informatics Conference (Poster), Denver, CO, May 18-21, 2026. 10 pages, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09691 2026-03-11 cs.CL cs.AI 79%

ESAinsTOD: A Unified End-to-End Schema-Aware Instruction-Tuning Framework for Task-Oriented Dialog Modeling

ESAinsTOD: 一种统一的端到端模式感知指令微调框架用于任务导向对话建模

Dechuan Teng, Chunlin Lu, Libo Qin, Wanxiang Che

机构 * Research Center for Social Computing and Information Retrieval, Harbin Institute of Technology(社会科学与信息检索研究中心,哈尔滨工业大学) School of Computer Science and Engineering, Central South University(计算机科学与工程学院,中南大学)

专题命中 指令微调 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 ESAinsTOD提出一种统一的端到端模式感知指令微调框架,提升任务导向对话建模的性能与泛化能力。

Comments Published at International Journal of Machine Learning and Cybernetics (IJMLC)

Journal ref Int. J. Mach. Learn. & Cyber. 17, 127 (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09758 2026-03-11 cs.CL 77%

Beyond Fine-Tuning: Robust Food Entity Linking under Ontology Drift with FoodOntoRAG

超越微调:在本体漂移下利用FoodOntoRAG实现鲁棒的食品实体链接

Jan Drole, Ana Gjorgjevikj, Barbara Korouši'c Seljak, Tome Eftimov

机构 * Computer Systems Department, Jožef Stefan International Postgraduate School, Ljubljana, Slovenia(卢布尔雅那,斯洛文尼亚乔泽夫·斯蒂芬国际研究生学院计算机系统系) Computer Systems Department, Jožef Stefan Institute, Ljubljana, Slovenia(卢布尔雅那,斯洛文尼亚乔泽夫·斯蒂芬研究所计算机系统系)

专题命中 指令微调 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出FoodOntoRAG,通过检索本体和结构化证据实现鲁棒的食品实体链接,避免微调并提高本体漂移的鲁棒性。

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09740 2026-03-11 cs.RO cs.CV 75%

Let's Reward Step-by-Step: Step-Aware Contrastive Alignment for Vision-Language Navigation in Continuous Environments

逐步奖励:面向连续环境的视觉语言导航中的步骤感知对比对齐

Haoyuan Li, Rui Liu, Hehe Fan, Yi Yang

机构 * College of Computer Science and Technology, Zhejiang University, Hangzhou, China(浙江大学计算机科学与技术学院,杭州,中国)

专题命中 指令微调 :large language model(abstract);language model(abstract);SFT(abstract)

AI总结 本文提出SACA框架,通过步骤感知对比对齐提升连续环境中视觉语言导航的性能,解决传统方法在错误恢复和训练稳定性方面的不足。

Comments 28 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09312 2026-03-11 cs.CV 75%

IntroSVG: Learning from Rendering Feedback for Text-to-SVG Generation via an Introspective Generator-Critic Framework

IntroSVG: 通过反思生成器-批评者框架从渲染反馈中学习以实现文本到SVG生成

Feiyu Wang, Jiayuan Yang, Zhiyuan Zhao, Da Zhang, Bingyu Li, Peng Liu, Junyu Gao

机构 * Fudan University(复旦大学) TeleAI Northwestern Polytechnical University(西北工业大学) University of Science and Technology of China(中国科学技术大学)

专题命中 指令微调 :language model(abstract);SFT(abstract);preference optimization(abstract)

AI总结 IntroSVG通过反思生成器-批评者框架,结合监督微调和直接偏好优化,实现文本到SVG生成的高质量输出。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09951 2026-03-11 cs.LG cs.AI cs.SE 73%

Towards a Neural Debugger for Python

迈向Python的神经调试器

Maximilian Beck, Jonas Gehring, Jannik Kossen, Gabriel Synnaeve

专题命中 指令微调 :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出神经调试器,通过模拟传统调试器实现交互式代码执行控制,提升神经模型在代码执行预测与逆向推理中的能力。

Comments 22 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.15447 2026-03-11 cs.LG cs.AI 62%

TSFM in-context learning for time-series classification of bearing-health status

基于时间序列基础模型的上下文学习用于滚动轴承健康状态分类

Michel Tokic, Slobodan Djukanović, Anja von Beuningen, Cheng Feng

机构 * Siemens AG Data & Artificial Intelligence(西门子AG数据与人工智能) Faculty of Mathematics Informatics and Statistics(数学信息统计系) Ludwig-Maximilians-University Munich(慕尼黑路德维希-马克西米利安大学) Faculty of Electrical Engineering(电气工程系) University of Montenegro(蒙特内格罗大学)

专题命中 指令微调 :foundation model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出基于时间序列基础模型的上下文学习方法,用于滚动轴承健康状态分类,通过伪时间序列模式预测类别概率,提升维护系统的智能化水平。

Comments Preprint. To appear in the Proceedings of the European Symposium on Artificial Neural Networks (ESANN), 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.08812 2026-03-11 cs.CV 50%

VisionCreator-R1: A Reflection-Enhanced Native Visual-Generation Agentic Model

VisionCreator-R1: 一种增强反思的原生视觉生成代理模型

Jinxiang Lai, Wenzhe Zhao, Zexin Lu, Hualei Zhang, Qinyu Yang, Rongwei Quan, Zhimin Li, Shuai Shao, Song Guo, Qinglin Lu

机构 * Tencent Hunyuan(腾讯文言) Hong Kong University of Science and Technology(香港科学与技术大学)

专题命中 指令微调 :SFT(abstract)

AI总结 VisionCreator-R1通过引入反射-计划联合优化方法,提升视觉生成代理在单图和多图任务中的表现,优于现有模型。

详情

展开后加载摘要…

URL PDF HTML 收藏