arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-03-10 至 2026-03-10 共收录 31 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 31 篇

2603.06836 2026-03-10 cs.CL cs.GL 89%

Validation of a Small Language Model for DSM-5 Substance Category Classification in Child Welfare Records

验证用于儿童福利记录DSM-5物质类别分类的小型语言模型

Brian E. Perron, Dragan Stoll, Bryan G. Victor, Zia Qia, Andreas Jud, Joseph P. Ryan

专题命中 其他LLM :language model(title,abstract);small language model(title);LLM(abstract);large language model(abstract)

AI总结 研究验证了本地部署的小型语言模型在儿童福利记录中对DSM-5物质类别进行多标签分类的可靠性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.16848 2026-03-10 cs.LG cs.AI 88%

Meta-RL Induces Exploration in Language Agents

元强化学习诱导语言智能体的探索

Yulun Jiang, Liangze Jiang, Damien Teney, Michael Moor, Maria Brbic

机构 * EPFL(苏黎世联邦理工学院) ETH Zurich(苏黎世联邦理工学院) Idiap Research Institute(Idiap研究机构)

专题命中 其他LLM :language agent(title,abstract);LLM(abstract);large language model(abstract);language model(abstract)

AI总结 LaMer通过元强化学习框架提升语言智能体的探索能力,显著提高多任务性能并增强泛化能力。

Comments ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.13606 2026-03-10 q-bio.NC cs.AI cs.CL cs.CV cs.LG 87%

LaVCa: LLM-assisted Visual Cortex Captioning

LaVCa: 基于大语言模型的视觉皮层描述生成

Takuya Matsuyama, Shinji Nishimoto, Yu Takagi

机构 * University of Osaka(大阪大学) National Institute of Information and Communications Technology(信息与通信技术国家研究所) Nagoya Institute of Technology(名古屋技术大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 LaVCa利用大语言模型生成视觉皮层体素选择性的详细描述,提升对大脑表示的理解。

Comments Accepted to ICLR 2026. Website: https://sites.google.com/view/lavca-llm/

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06842 2026-03-10 cs.RO 85%

RoboCritics: Enabling Reliable End-to-End LLM Robot Programming through Expert-Informed Critics

RoboCritics: 通过专家指导的批评者实现可靠的端到端LLM机器人编程

Callie Y. Kim, Nathan Thomas White, Evan He, Frederic Sala, Bilge Mutlu

机构 * Department of Computer Sciences University of Wisconsin--Madison(计算机科学系威斯康星大学麦迪逊分校)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract)

AI总结 RoboCritics通过专家指导的批评者提升LLM在机器人编程中的可靠性与用户参与度。

Comments 10 pages, 5 figures, Proceedings of the 21st ACM/IEEE International Conference on Human Robot Interaction (HRI 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.07736 2026-03-10 econ.TH 83%

Menu Pricing of Large Language Models

大型语言模型的菜单定价

Dirk Bergemann, Alessandro Bonatti, Alex Smolin

专题命中 其他LLM :large language model(title);language model(title)

AI总结 本文提出了一种针对大型语言模型的最优定价框架,通过一维筛选机制和承诺支出合同,优化用户预算分配与计算资源提供。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.22989 2026-03-10 cs.AI cs.CY cs.GT 83%

Towards Strategic Persuasion with Language Models

面向语言模型的战略说服

Zirui Cheng, Jiaxuan You

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.AI

AI总结 本文基于贝叶斯说服理论,利用人类说服数据集构建环境,通过强化学习训练语言模型进行战略说服,验证了LLMs在不同规模下的说服能力

Comments ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.08520 2026-03-10 cs.CR cs.SE 82%

SCAFFOLD-CEGIS: Preventing Latent Security Degradation in LLM-Driven Iterative Code Refinement

SCAFFOLD-CEGIS: 防止由LLM驱动的迭代代码精炼中的潜在安全退化

Yi Chen, Yun Bian, Haiquan Wang, Shihao Li, Zhe Cui

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract)

AI总结 SCAFFOLD-CEGIS通过多智能体协作架构,将安全约束显式化,有效降低LLM驱动迭代代码精炼中的潜在安全退化率至2.1%。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00969 2026-03-10 cs.AI cs.SY eess.SY 79%

Integrating a Causal Foundation Model into a Prescriptive Maintenance Framework for Optimising Production-Line OEE

将因果基础模型整合到指令性维护框架中以优化生产线OEE

Felix Saretzky, Lucas Andersen, Thomas Engel, Fazel Ansari

机构 * Department of Engineering University of Luxembourg(工程系卢森堡大学) Department of Computer Science University of Luxembourg(计算机科学系卢森堡大学) Chair of Production and Maintenance Management TU Wien(生产与维护管理系维也纳技术大学)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI

AI总结 本文提出基于因果机器学习的模型,通过模拟潜在修复方案优化生产线OEE,解决传统预测模型无法识别故障根本原因的问题。

Comments 9 pages, 3 images, 1 table, conference paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.07717 2026-03-10 cs.AI cs.GT cs.HC 79%

Rigidity in LLM Bandits with Implications for Human-AI Dyads

在LLM老虎机中的刚性及其对人机双元体的影响

Haomiaomiao Wang, Tomás E Ward, Lili Zhang

机构 * Insight Research Ireland Centre for Data Analytics, Ireland(爱尔兰洞察研究爱尔兰数据分析中心) School of Computing, Dublin City University, Ireland(都柏林城市大学计算机学院)

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI

AI总结 研究发现LLM在老虎机任务中表现出刚性决策策略,通过计算建模揭示了低学习率和高逆温度的机制,为理解人机交互中的决策偏见提供了新视角。

Comments 13 pages, 5 figures, AICS conference https://aicsconf.org/

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.07766 2026-03-10 cs.CL cs.AI 79%

QuadAI at SemEval-2026 Task 3: Ensemble Learning of Hybrid RoBERTa and LLMs for Dimensional Aspect-Based Sentiment Analysis

QuadAI在SemEval-2026任务3中的表现:混合RoBERTa与LLMs的集成学习用于维度方面基于情感分析

A. J. W. de Vink, Filippos Karolos Ventirozos, Natalia Amat-Lefort, Lifeng Han

机构 * LIACS, Leiden University, NL(莱顿大学LIACS研究中心) Manchester Metropolitan University(曼彻斯特 Metropolitan 大学) Leiden University Medical Center(莱顿大学医学中心)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 QuadAI通过集成混合RoBERTa与LLMs的方法,在SemEval-2026任务3中实现了维度方面基于情感分析的高性能表现。

Comments SemEval System Report

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.02428 2026-03-10 cs.HC 78%

Can LLM-Simulated Practice and Feedback Upskill Human Counselors? A Randomized Study with 90+ Novice Counselors

大语言模型模拟实践与反馈能否提升人类辅导员能力?一项涉及90多名初级辅导员的随机研究

Ryan Louie, Raj Sanjay Shah, Ifdita Hasan Orney, Juan Pablo Pacheco, Emma Brunskill, Diyi Yang

专题命中 其他LLM :LLM(title,abstract)

AI总结 本研究通过随机对照试验发现,大语言模型模拟实践结合反馈能有效提升初级辅导员的客户中心微技能和共情能力,而单纯实践效果有限。

Comments Two-column conference proceedings format is 19 pages, with references and appendix it is 31 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.07442 2026-03-10 cs.RO 75%

LITHE: Bridging Best-Effort Python and Real-Time C++ for Hot-Swapping Robotic Control Laws on Commodity Linux

LITHE:连接最佳努力Python与实时C++以实现机器人控制律的热插拔

He Kai Lim, Tyler R. Clites

机构 * Department of Mechanical and Aerospace Engineering, University of California Los Angeles(加州大学洛杉矶分校机械与航空航天工程系)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 LITHE通过轻量级架构实现Python与C++的实时控制律热插拔,提升机器人系统在动态环境中的适应能力。

Comments 8 pages, 5 figures. Submitted to IEEE/RSJ International Conference on Intelligent Robots & Systems (IROS) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06647 2026-03-10 cs.NI cs.AI 74%

Performance Comparison of IBN orchestration using LLM and SLMs

基于LLM和SLMs的IBN编排性能比较

Wai Lwin Phone, Brahim El Boudani, Tasos Dagiuklas, Saptarshi Ghosh

机构 * 1 Dept. of Computer Science \& Digital Technologies, London South Bank University, London, UK

专题命中 其他LLM :LLM(title);分类 cs.AI

AI总结 本文比较了LLM和SLMs在IBN编排中的性能,发现SLMs能提升IBN生命周期完成速度20%。

Comments Accepted for presentation at IEEE International Conference on Communications 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.18907 2026-03-10 cs.AI cs.LG 73%

Stronger Enforcement of Instruction Hierarchy via Augmented Intermediate Representations

通过增强的中间表示实现更强的指令层级强制执行

Sanjay Kariyappa, G. Edward Suh

机构 * NVIDIA(英伟达)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 通过在中间表示中注入增强的指令层级信号,有效降低提示注入攻击的成功率,同时保持模型性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.08604 2026-03-10 cs.HC 71%

What to Make Sense of in the Era of LLM? A Perspective from the Structure and Efforts in Sensemaking

在大语言模型时代如何寻求意义?从意义建构的结构和努力角度的视角

Tianyi Li, Satya Samhita Bonepalli, Vikram Mohanty

专题命中 其他LLM :LLM(title)

AI总结 本文探讨了GPT-4在解读虚构恐怖分子计划中的意义建构任务,通过整体和逐步两种方法,探索人机协作提升复杂情境下的意义建构效率。

Comments CHI 2024 Sensemaking Workshop https://sites.google.com/view/chi2024-sensemaking-workshop/home?pli=1

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.07817 2026-03-10 cs.CV 71%

Tracking Phenological Status and Ecological Interactions in a Hawaiian Cloud Forest Understory using Low-Cost Camera Traps and Visual Foundation Models

利用低成本相机陷阱和视觉基础模型追踪夏威夷云林 understory 的物候状态和生态互动

Luke Meyers, Anirudh Potlapally, Yuyan Chen, Mike Long, Tanya Berger-Wolf, Hari Subramoni, Remi Megret, Daniel Rubenstein

机构 * Computer Science, The Ohio State University, Columbus, Ohio, USA(俄亥俄州立大学计算机科学系) Computer Science, The University of Puerto Rico Rio Piedras, San Juan, Puerto Rico(波多黎各大学里奥皮埃拉斯分校计算机科学系) Computer Science, McGill University, Montreal, Quebec, Canada(麦吉尔大学计算机科学系) Battele Ecology, Neon Domain 20, Hilo, Hawaii(Battele生态研究所) Ecology, Princeton University, Princeton, New Jersey, USA(普林斯顿大学生态学系)

专题命中 其他LLM :foundation model(title)

AI总结 本研究利用低成本相机陷阱和视觉基础模型,追踪夏威夷云林 understory 的植物物候变化及生态互动,揭示时间精细度高的物候趋势。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19112 2026-03-10 cs.CV 67%

Universal 3D Shape Matching via Coarse-to-Fine Language Guidance

通过粗到细的语言引导实现通用3D形状匹配

Qinfeng Xiao, Guofeng Mei, Bo Yang, Liying Zhang, Jian Zhang, Kit-lun Yick

机构 * Hong Kong Polytechnic University, HK SAR(香港理工大学) Fondazione Bruno Kessler, Italy(布鲁诺·凯斯勒基金会) University of Technology Sydney, Australia(悉尼科技大学)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 UniMatch通过粗到细的语言引导方法,实现跨类别非等距形状的通用3D匹配。

Comments Accepted by CVPR 2026

Journal ref CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.07155 2026-03-10 cs.HC cs.MA 67%

NarrativeLoom: Enhancing Creative Storytelling through Multi-Persona Collaborative Improvisation

NarrativeLoom:通过多角色协作即兴创作增强创造性叙事

Yuxi Ma, Yongqian Peng, Fengyuan Yang, Siyu Zha, Chi Zhang, Zixia Jia, Zilong Zheng, Yixin Zhu

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 NarrativeLoom通过多角色协作即兴创作提升叙事原创性,通过理论指导的协作系统增强创造性输出。

Comments 19 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.07134 2026-03-10 cs.HC 67%

More Than 1v1: Human-AI Alignment in Early Developmental Communities with Multimodal LLMs

多于一对一:在早期发展社区中的人工智能对齐与多模态大语言模型

Weiyan Shi, Kenny Tsu Wei Choo

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文探讨了在早期发展社区中,利用多模态大语言模型实现人机对齐的挑战,提出分层社区对齐框架,强调社区治理而非个体优化。

Comments Accepted at CHI 2026 BiAlign Workshop; OpenReview URL: https://openreview.net/forum?id=ikeH0hsBLN

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.08572 2026-03-10 cs.RO cs.AI 57%

MetaWorld-X: Hierarchical World Modeling via VLM-Orchestrated Experts for Humanoid Loco-Manipulation

MetaWorld-X: 通过VLM协调的专家实现人形机器人的分层世界建模

Yutong Shen, Hangxu Liu, Penghui Liu, Jiashuo Luo, Yongkang Zhang, Rex Morvley, Chen Jiang, Jianwei Zhang, Lei Zhang

机构 * University of Hamburg(汉堡大学) School of Information Science and Technology, Beijing University of Technology(信息科学与技术学院,北京理工大学) School of Information Science and Engineering, Fudan University(信息科学与工程学院,复旦大学) University of Alberta(阿尔伯塔大学)

专题命中 其他LLM :language model(abstract);分类 cs.AI

AI总结 MetaWorld-X通过VLM协调的专家实现人形机器人分层世界建模,解决loco-manipulation任务中的控制策略泛化问题。

Comments 8 figures, https://syt2004.github.io/metaworldX/

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05437 2026-03-10 cs.CV cs.AI 57%

SAIL: Similarity-Aware Guidance and Inter-Caption Augmentation-based Learning for Weakly-Supervised Dense Video Captioning

SAIL:基于相似性感知引导和跨字幕增强学习的弱监督密集视频字幕生成

Ye-Chan Kim, SeungJu Cha, Si-Woo Kim, Minju Jeon, Hyungee Kim, Dong-Jin Kim

机构 * Hanyang University(翰阳大学)

专题命中 其他LLM :LLM(abstract);分类 cs.AI

AI总结 SAIL通过跨模态对齐和大语言模型增强策略,提升弱监督密集视频字幕生成的准确性和语义感知能力。

Comments Accepted to CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.22519 2026-03-10 cs.AI cs.IT math.IT 57%

A Mathematical Theory of Agency and Intelligence

智能与代理的数学理论

Wael Hafez, Chenan Wei, Rodrigo Pena, Amir Nazeri, Cameron Reid

专题命中 其他LLM :LLM(abstract);分类 cs.AI

AI总结 本文提出双可预测性P作为衡量智能与代理的核心指标,区分了代理与智能,并展示了反馈架构对实现适应性AI的重要性。

Comments 20 pages, 4 figuers

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00126 2026-03-10 q-bio.QM cs.AI 57%

RadDiff: Retrieval-Augmented Denoising Diffusion for Protein Inverse Folding

RadDiff:基于检索的去噪扩散用于蛋白质逆折叠

Jin Han, Tianfan Fu, Wu-Jun Li

机构 * National Key Laboratory for Novel Software Technology(国家新型软件技术重点实验室) School of Computer Science, Nanjing University(南京大学计算机科学学院)

专题命中 其他LLM :language model(abstract);分类 cs.AI

AI总结 RadDiff通过引入检索增强机制和知识感知扩散模型,提升蛋白质逆折叠的序列恢复率和可折叠性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.07246 2026-03-10 cs.CV cs.AI 57%

LEPA: Learning Geometric Equivariance in Satellite Remote Sensing Data with a Predictive Architecture

LEPA: 在卫星遥感数据中学习几何等变性以预测架构

Erik Scheurer, Rocco Sedona, Stefan Kesselheim, Gabriele Cavallaro

机构 * Jülich Supercomputing Centre (JSC), Forschungszentrum Jülich(茹里希超级计算中心(JSC),茹里希研究中心) Institute for Visualization and Interactive Systems (VIS), University of Stuttgart(可视化与交互系统研究所(VIS),斯图加特大学) School of Engineering and Natural Sciences (SENS), University of Iceland(工程与自然科学学院(SENS),爱沙尼亚大学)

专题命中 其他LLM :foundation model(abstract);分类 cs.AI

AI总结 LEPA通过学习几何等变性预测架构,提升卫星遥感数据中几何调整的准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17788 2026-03-10 cs.CV cs.AI 57%

From 2D Alignment to 3D Plausibility: Unifying Heterogeneous 2D Priors and Penetration-Free Diffusion for Occlusion-Robust Two-Hand Reconstruction

从二维对齐到三维合理性:统一异构二维先验和无穿透扩散以实现抗遮挡的双手重建

Gaoge Han, Yongkang Cheng, Zhe Chen, Shaoli Huang, Tongliang Liu

机构 * AgiBot Mohamed bin Zayed University of Artificial Intelligence The University of Sydney La Trobe University

专题命中 其他LLM :foundation model(abstract);分类 cs.AI

AI总结 本文提出统一异构二维先验和无穿透扩散模型,以实现抗遮挡的双手重建,提升交互对齐和穿透抑制性能。

Comments Accepted by CVPR 2026 Main, Project: https://gaogehan.github.io/A2P/

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.08616 2026-03-10 cs.SE cs.CR 50%

Coverage-Guided Multi-Agent Harness Generation for Java Library Fuzzing

面向 Java 库的覆盖引导多智能体模糊生成

Nils Loose, Nico Winkel, Kristoffer Hempel, Felix Mächtle, Julian Hans, Thomas Eisenbarth

专题命中 其他LLM :LLM(abstract)

AI总结 本文提出一种多智能体架构,通过 LLM 功能智能体自动化生成 Java 库的模糊 harnesses,提升模糊测试效率并发现漏洞。

Comments Accepted at The 19th International Workshop on Search-Based and Fuzz Testing (SBFT 2026, ICSE Workshop)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.08532 2026-03-10 physics.optics 50%

Recent advances in spatial light modulator-based three-dimensional optical imaging (Invited)

基于空间光调制器的三维光学成像近期进展(特邀)

Joseph Rosen

专题命中 其他LLM :SLM(abstract)

AI总结 本文综述了基于空间光调制器的三维光学成像技术,涵盖从传统二维成像到多维成像、轴向切片等应用,展示了该领域近年来的重要进展。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00481 2026-03-10 cs.HC 50%

From Performers to Creators: Understanding Retired Women's Perceptions of Technology-Enhanced Dance Performance

从表演者到创作者:理解退休女性对技术增强舞蹈表演的认知

Danlin Zheng, Xiaoying Wei, Chao Liu, Quanyu Zhang, Jingling Zhang, Shihui Guo, Mingming Fan

专题命中 其他LLM :LLM(abstract)

AI总结 本文通过交互式舞蹈技术与AI生成内容,帮助退休女性克服年龄相关限制,提升舞台表现并成为表演共创者。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.07028 2026-03-10 cs.CR cs.CV 50%

Two Frames Matter: A Temporal Attack for Text-to-Video Model Jailbreaking

两个框架至关重要:一种针对文本到视频模型劫持的时序攻击

Moyang Chen, Zonghao Ying, Wenzhuo Xu, Quancheng Zou, Deyue Zhang, Dongdong Yang, Xiangzheng Zhang

机构 * College of Science, Mathematics and Technology, Wenzhou-Kean University(科学、数学与技术学院,温州市凯恩大学) State Key Laboratory of Complex & Critical Software Environment, Beihang University(复杂与关键软件环境国家重点实验室,北航) AI Security Lab(360人工智能安全实验室)

专题命中 其他LLM :prompting(abstract)

AI总结 本文提出TFM攻击方法,通过碎片化提示和隐式替换,提高文本到视频模型的劫持效果,显著提升攻击成功率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06756 2026-03-10 q-bio.QM 50%

GWAS Summary Statistic Tool: A Meta-Analysis and Parsing Tool for Polygenic Risk Score Calculation

GWAS总结统计工具:用于多基因风险评分计算的元分析和解析工具

Muhammad Muneeb, David B. Ascher

专题命中 其他LLM :LLM(abstract)

AI总结 GWASPoker是一种用于多基因风险评分计算的GWAS总结统计工具,通过部分下载和标题检测自动筛选和解析GWAS文件,提高PRS计算效率。

详情

展开后加载摘要…

URL PDF HTML 收藏