arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-03-03 至 2026-03-03 共收录 33 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 33 篇

2602.04369 2026-03-03 cs.LG 89%

Multi-scale hypergraph meets LLMs: Aligning large language models for time series analysis

多尺度超图与大语言模型:面向时间序列分析的对齐方法

Zongjiang Shang, Dongliang Cui, Binqing Wu, Ling Chen

机构 * State Key Laboratory of Blockchain and Data Security, Zhejiang University(区块链与数据安全国家重点实验室,浙江大学) College of Computer Science and Technology, Zhejiang University(计算机科学与技术学院,浙江大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.LG

AI总结 本文提出MSH-LLM方法,通过多尺度超图机制和跨模态对齐模块,提升大语言模型在时间序列分析中的表现。

Comments Accepted by ICLR2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01239 2026-03-03 cs.CL cs.AI 88%

Self-Anchoring Calibration Drift in Large Language Models: How Multi-Turn Conversations Reshape Model Confidence

大语言模型中的自我锚定校准漂移:多轮对话如何重塑模型信心

Harshavardhan

机构 * Independent Researcher(独立研究者)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 研究发现大语言模型在多轮对话中出现自我锚定校准漂移,导致模型信心变化及校准误差波动。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.21895 2026-03-03 cs.CL cs.AI stat.ML 86%

Learn-to-Distance: Distance Learning for Detecting LLM-Generated Text

学以致用:用于检测LLM生成文本的距离学习

Hongyi Zhou, Jin Zhu, Kai Ye, Ying Yang, Erhan Xu, Chengchun Shi

机构 * Tsinghua University(清华大学) University of Birmingham(伯明翰大学) London School of Economics and Political Science(伦敦政治经济学院)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本文提出了一种自适应学习的距离方法,用于更有效地检测LLM生成的文本,通过几何方法揭示了重写检测算法的原理,并在多种LLM上实现了显著的性能提升。

Comments Accepted by ICLR2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01942 2026-03-03 cs.HC cs.AI 85%

Ignore All Previous Instructions: Jailbreaking as a de-escalatory peace building practise to resist LLM social media bots

忽略此前所有指令:将‘禁锢’视为一种缓和和平建设的实践,以抵抗大语言模型社交媒体机器人

Huw Day, Adrianna Jezierska, Jessica Woodgate

机构 * School of Engineering Maths & Technology University of Bristol(工程数学与科技学院 英国布里斯托尔大学) Business School University of Bristol(商学院 英国布里斯托尔大学) School of Computer Science University of Bristol(计算机科学学院 英国布里斯托尔大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出将‘禁锢’作为一种非暴力的缓和和平建设实践,通过用户与疑似LLM账号的互动来对抗大语言模型社交媒体机器人。

Comments Accepted to ICLR 2026 AI for peace workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01638 2026-03-03 cs.HC 85%

Who Explains Privacy Policies to Me? Embodied and Textual LLM-Powered Privacy Assistants in Virtual Reality

谁向我解释隐私政策?基于大型语言模型的虚拟现实隐私助手

Vincent Freiberger, Moritz Dresch, Florian Alt, Arthur Fleig, Viktorija Paneva

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract)

AI总结 本研究提出基于大型语言模型的虚拟现实隐私助手,通过文本和具身交互模式提升用户对隐私政策的理解和决策质量。

Comments 11 pages, 1 figure, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00669 2026-03-03 cs.CL cs.AI cs.HC 84%

SSKG Hub: An Expert-Guided Platform for LLM-Empowered Sustainability Standards Knowledge Graphs

SSKG Hub: 一个基于大语言模型的可持续性标准知识图谱专家指导平台

Chaoyue He, Xin Zhou, Xinjia Yu, Lei Zhang, Yan Zhang, Yi Wu, Lei Xiao, Liangyue Li, Di Wang, Hong Xu, Xiaoqiao Wang, Wei Liu, Chunyan Miao

机构 * Alibaba-NTU Global e-Sustainability CorpLab (ANGEL)(阿里-国立大学全球可持续性公司实验室(ANGEL)) Alibaba Group(阿里巴巴集团)

专题命中 其他LLM :LLM(title,abstract);prompting(abstract);分类 cs.CL、cs.AI

AI总结 SSKG Hub通过大语言模型和专家指导构建可持续性标准知识图谱,实现标准到可审计图谱的转化,并提供治理框架和跨图谱融合功能。

Comments 10 pages, 2 figures, 2 tables, submitted to ACL26 System Demo Track

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.22210 2026-03-03 cs.SE cs.AI 81%

LSPRAG: LSP-Guided RAG for Language-Agnostic Real-Time Unit Test Generation

LSPRAG: 语言无关的实时单元测试生成中的LSP引导RAG

Gwihwan Go, Quan Zhang, Chijin Zhou, Zhao Wei, Yu Jiang

机构 * Tsinghua University(清华大学) East China Normal University(华东师范大学) Tencent(腾讯)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

AI总结 LSPRAG通过利用语言服务器协议实现实时语言无关单元测试生成,显著提高了测试覆盖率。

Comments 13pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01254 2026-03-03 cs.CL cs.AI 81%

LLM Self-Explanations Fail Semantic Invariance

LLM 自我解释失效语义不变性

Stefan Szeider

机构 * Algorithms and Complexity Group TU Wien(算法与复杂性组维也纳技术大学)

专题命中 其他LLM :LLM(title,abstract);分类 cs.CL、cs.AI

AI总结 LLM自我解释在语义上下文变化时无法保持稳定,表明其无法可靠反映模型内部状态或能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00522 2026-03-03 cs.HC 80%

SIAgent: Spatial Interaction Agent via LLM-powered Eye-Hand Motion Intent Understanding in VR

SIAgent:通过LLM驱动的眼手运动意图理解实现VR中的空间交互代理

Zhimin Wang, Chenyu Gu, Feng Lu

专题命中 其他LLM :LLM(title,abstract);large language model(comments);language model(comments)

AI总结 SIAgent通过LLM驱动的意图识别与代理执行,实现VR中基于自然眼手运动的高容忍度交互框架。

Comments Virtual reality, spatial interaction, intent recognition, agent-based execution, large language models

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01824 2026-03-03 cs.CL cs.LG 79%

OpenAutoNLU: Open Source AutoML Library for NLU

OpenAutoNLU: 开源自然语言理解自动机器学习库

Grigory Arshinov, Aleksandr Boriskin, Sergey Senichev, Ayaz Zaripov, Daria Galimzianova, Daniil Karpov, Leonid Sanochkin

机构 * MWS AI ITMO University(ITMO大学) MBZUAI

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

AI总结 OpenAutoNLU是一个开源的自动机器学习库,提供自然语言理解任务的文本分类和命名实体识别功能,通过数据感知的训练制度选择和低代码API实现高效自动化。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01596 2026-03-03 cs.SE 78%

MigMate: A VS Code Extension for LLM-based Library Migration of Python Projects

MigMate: 一个用于Python项目基于LLM的库迁移的VS Code扩展

Matthias Kebede, May Mahmoud, Mohayeminul Islam, Sarah Nadi

专题命中 其他LLM :LLM(title,abstract)

AI总结 MigMate是一款基于LLM的VS Code插件,通过集成自动化迁移过程,帮助开发者更高效地完成Python项目库迁移任务。

Comments 6 pages, 6 figures, 2 tables, 3rd International Workshop on Integrated Development Environments (IDE 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01773 2026-03-03 cs.CL 77%

AnnoABSA: A Web-Based Annotation Tool for Aspect-Based Sentiment Analysis with Retrieval-Augmented Suggestions

AnnoABSA:一个支持基于方面的情感分析的网页标注工具,具有检索增强的建议

Nils Constantin Hellwig, Jakob Fehle, Udo Kruschwitz, Christian Wolff

机构 * Media Informatics Group, University of Regensburg, Regensburg, Germany(里根斯堡大学媒体信息学组) Information Science Group, University of Regensburg, Regensburg, Germany(里根斯堡大学信息科学组)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 AnnoABSA是一款支持基于方面的情感分析的网页标注工具,结合检索增强生成建议提升标注效率与准确性。

Comments Accepted for publication at LREC 2026. Final version will appear in the ACL Anthology

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00003 2026-03-03 cs.CY cs.CL cs.DL 77%

Commitment Checklist: Auditing Author Commitments in Peer Review

承诺清单:在同行评审中审计作者承诺

Chung-Chi Chen, Iryna Gurevych

机构 * AIST, Japan(日本国家信息技术研究所) Ubiquitous Knowledge Processing Lab (UKP Lab), Department of Computer Science, TU Darmstadt and National Research Center for Applied Cybersecurity ATHENE, Germany(图宾根大学计算机科学系通用知识处理实验室及应用网络安全国家研究中心ATHENE德国)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出作者承诺清单,利用LLM审计同行评审中作者的承诺,发现约25%的承诺未被履行,强调了加强同行评审问责制的重要性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.10067 2026-03-03 cs.CV 75%

FiLo++: Zero-/Few-Shot Anomaly Detection by Fused Fine-Grained Descriptions and Deformable Localization

FiLo++: 通过融合细粒度描述和变形定位实现零/少样本异常检测

Zhaopeng Gu, Bingke Zhu, Guibo Zhu, Yingying Chen, Ming Tang, Jinqiao Wang

机构 * Foundation Model Research Center, Institute of Automation, Chinese Academy of Sciences(基础模型研究中心,自动化研究所,中国科学院) School of Artifcial Intelligence, University of Chinese Academy of Sciences(人工智能学院,中国科学院大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) Wuhan AI Research(武汉AI研究所) Peng Cheng Laboratory(鹏城实验室) Guangdong Provincial Key Laboratory of Intellectual Property & Big Data, Guangdong Polytechnic Normal University(广东省知识产权与大数据重点实验室,广东工业大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);foundation model(abstract)

AI总结 FiLo++通过融合细粒度描述和变形定位技术,提升零/少样本异常检测的准确性和鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.15863 2026-03-03 cs.CL cs.AI 73%

PolySkill: Learning Generalizable Skills Through Polymorphic Abstraction

PolySkill: 通过多态抽象学习通用技能

Simon Yu, Gang Li, Weiyan Shi, Peng Qi

机构 * Northeastern University(东北大学) Uniphore

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 PolySkill通过多态抽象方法使智能体能学习通用且可组合的技能,提升技能重用性和成功率,减少步骤并增强自我探索能力。

Comments 29 pages, 6 figures, 8 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21910 2026-03-03 cs.LG 70%

Adversarial Déjà Vu: Jailbreak Dictionary Learning for Stronger Generalization to Unseen Attacks

对抗性 déjà vu:用于更强泛化到未见攻击的 jailbreak 字典学习

Mahavir Dabas, Tran Huynh, Nikhil Reddy Billa, Jiachen T. Wang, Peng Gao, Charith Peris, Yao Ma, Rahul Gupta, Ming Jin, Prateek Mittal, Ruoxi Jia

机构 * Virginia Tech(弗吉尼亚理工大学) Princeton University(普林斯顿大学) Amazon AGI(亚马逊人工智能研究院)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.LG

AI总结 本文提出通过对抗性技能组合训练提升模型对未知攻击的鲁棒性,通过分析攻击论文发现新型攻击是旧技能的重新组合,从而改进安全防御。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.19025 2026-03-03 cs.DB cs.CL 70%

SQUiD: Synthesizing Relational Databases from Unstructured Text

从无结构文本合成关系数据库

Mushtari Sadia, Zhenning Yang, Yunming Xiao, Ang Chen, Amrita Roy Chowdhury

机构 * University of Michigan(密歇根大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 SQUiD通过神经符号框架从无结构文本自动合成关系数据库,优于现有方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.17489 2026-03-03 cs.HC cs.AI 70%

Neural Spelling: A Spell-Based BCI System for Language Neural Decoding

神经拼写:一种基于拼写的BCI系统用于语言神经解码

Xiaowei Jiang, Charles Zhou, Yiqun Duan, Ziyi Zhao, Thomas Do, Chin-Teng Lin

机构 * Brain Computer Interface Lab, School of Computer Science, Faculty of Engineering Information Technology, University of Technology Sydney

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出一种基于EEG的BCI系统,结合课程神经拼写框架和生成式AI,实现高准确率的字母识别和文本生成,提升残障人士的交流能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01505 2026-03-03 cs.RO 67%

FATE: Closed-Loop Feasibility-Aware Task Generation with Active Repair for Physically Grounded Robotic Curricula

FATE: 闭环可行性感知任务生成与主动修复的物理 grounded 机器人课程

Bingchuan Wei, Bingqi Huang, Jingheng Ma, Zeyu zhang, Sen Cui

机构 * School of Aerospace Engineering, Tsinghua University, Beijing, China(航空航天工程系,清华大学,北京,中国) Department of Automation, Tsinghua University, Beijing, China(自动化系,清华大学,北京,中国) School of Integrated Circuits, Tsinghua University, Beijing, China(集成电路学院,清华大学,北京,中国) State Key Laboratory of General Artificial Intelligence, Beijing Institute for General Artificial Intelligence (BIGAI), Beijing, China(通用人工智能国家重点实验室,北京通用人工智能研究院(BIGAI),北京,中国)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 FATE通过闭环验证与主动修复机制,生成物理 grounded 的机器人任务课程,有效减少执行失败率。

Comments 16 Pages, 4 Figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01319 2026-03-03 cs.HC 67%

Caught in a Mafia Romance: How Users Explore Intimate Roleplay and Narrative Exploration with Chatbots

陷入黑手党浪漫:用户如何通过聊天机器人探索亲密角色扮演与叙事探索

Julia Kieserman, Cat Mai, Sara Lignell, Lucy Qin, Athanasios Andreou, Damon McCoy, Rosanna Bellini

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 研究探讨用户通过聊天机器人进行亲密角色扮演和幻想探索的行为,发现用户偏好特定角色设定并对其内容的性化程度提出安全需求。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01314 2026-03-03 cs.HC 67%

Actor's Note: Examining the Role of AI-Generated Questions in Character Journaling for Actor Training

演员笔记:探讨AI生成问题在演员训练中的角色期刊作用

Sora Kang, Jaemin Zoh, Hyoju Kim, Hyeonseo Park, Hajin Lim, Joonhwan Lee

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 Actor's Note通过AI生成问题辅助演员训练,提升角色探索与反思实践,保持艺术沉浸感。

Comments In Proceedings of the 2026 CHI Conference on Human Factors in Computing Systems (CHI '26), April 13-17, 2026, Barcelona, Spain

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.22171 2026-03-03 cs.HC 67%

A Taxonomy of Human--MLLM Interaction in Early-Stage Sketch-Based Design Ideation

早期阶段基于草图的设计构想中人类与大语言模型交互的分类

Weiyan Shi, Kenny Tsu Wei Choo

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出了一种分类方法,用于描述人类与大语言模型在早期阶段基于草图的设计构想中的交互模式,揭示了人类与AI角色的动态变化。

Comments Accepted at CHI 2026 Posters

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01022 2026-03-03 cs.CE 67%

GeoMCP: A Trustworthy Framework for AI-Assisted Analytical Geotechnical Engineering

GeoMCP:一种可信的AI辅助分析土木工程框架

Yared W. Bekele

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 GeoMCP通过将工程方法表示为结构化数据,构建了一个可信的AI辅助分析土木工程框架,确保计算透明性和安全性。

Comments 15 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01694 2026-03-03 cs.CV cs.AI cs.LG 62%

MVR: Multi-view Video Reward Shaping for Reinforcement Learning

MVR:多视图视频奖励塑造用于强化学习

Lirui Luo, Guoxi Zhang, Hongming Xu, Yaodong Yang, Cong Fang, Qing Li

机构 * School of Intelligence Science and Technology, Peking University(北京大学智能科学与技术学院) State Key Laboratory of General Artificial Intelligence, BIGAI(通用人工智能国家重点实验室)

专题命中 其他LLM :language model(abstract);分类 cs.AI、cs.LG

AI总结 MVR通过多视角视频和视觉语言模型提升强化学习中的奖励塑造,有效解决复杂动态任务中的状态相关性和视角偏见问题。

Comments ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01339 2026-03-03 stat.ML cs.LG 57%

Causal Effects with Unobserved Unit Types in Interacting Human-AI Systems

交互人类-人工智能系统中未观察到的单元类型因果效应

William Overman, Sadegh Shirani, Mohsen Bayati

专题命中 其他LLM :LLM(abstract);分类 cs.LG

AI总结 研究提出在未观察单元类型和交互网络的情况下,通过因果信息传递框架估计人类特定因果效应的方法,并在模拟平台验证其有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01048 2026-03-03 cs.SE cs.AI 57%

RepoRepair: Leveraging Code Documentation for Repository-Level Automated Program Repair

RepoRepair: 利用代码文档实现仓库级别的自动程序修复

Zhongqiang Pan, Chuanyi Li, Wenkang Zhong, Yi Feng, Bin Luo, Vincent Ng

机构 * State Key Laboratory for Novel Software Technology, Nanjing University(新型软件技术国家重点实验室,南京大学) Human Language Technology Research Institute, University of Texas at Dallas(人机语言技术研究院,德克萨斯大学达拉斯分校)

专题命中 其他LLM :LLM(abstract);分类 cs.AI

AI总结 RepoRepair通过生成代码文档增强LLM能力,实现仓库级别的自动程序修复,取得高修复率和低成本的优异表现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00529 2026-03-03 cs.CV cs.AI 57%

CaptionFool: Universal Image Captioning Model Attacks

CaptionFool: 针对最新Transformer图像描述模型的通用图像描述模型攻击

Swapnil Parekh

机构 * Intuit

专题命中 其他LLM :language model(abstract);分类 cs.AI

AI总结 CaptionFool通过修改少量图像块,成功生成任意目标描述,揭示了视觉-语言模型的安全漏洞。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.09085 2026-03-03 cs.HC cs.AI cs.CY 57%

Mental Models of Autonomy and Sentience Shape Reactions to AI

自主性与意识的内心模型影响对AI的反应

Janet V. T. Pauketat, Daniel B. Shank, Aikaterina Manoli, Jacy Reese Anthis

机构 * Sentience Institute(意识研究所) Missouri University of Science and Technology(密苏里科技大学) Max Planck Institute for Human Cognitive and Brain Sciences(人类认知与脑科学Max Planck研究所) Stanford University(斯坦福大学) University of Chicago(芝加哥大学)

专题命中 其他LLM :prompting(abstract);分类 cs.AI

AI总结 研究探讨自主性与意识的内心模型如何影响人类对AI的反应,发现意识比自主性更能引发道德考虑,而自主性则增加威胁感知。

Comments Published at CHI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01415 2026-03-03 eess.AS 50%

The USTC-NERCSLIP Systems for the CHiME-9 MCoRec Challenge

USTC-NERCSLIP系统参加CHiME-9 MCoRec挑战

Ya Jiang, Ruoyu Wang, Jingxuan Zhang, Jun Du, Yi Han, Zihao Quan, Hang Chen, Yeran Yang, Kongzhi Zheng, Zhuo Chen, Yanhui Tu, Shutong Niu, Changfeng Xi, Mengzhi Wang, Zhongbin Wu, Jieru Chen, Henghui Zhi, Weiyi Shi, Shuhang Wu, Genshun Wan, Jia Pan, Jianqing Gao

专题命中 其他LLM :LLM(abstract)

AI总结 USTC-NERCSLIP系统通过多模态级联方法和LLM技术,在CHiME-9 MCoRec挑战中实现15.70%的联合ASR-聚类错误率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00722 2026-03-03 econ.GN q-fin.EC 50%

On Repeat: Does Iteration Drive Innovation?

重复与创新:迭代是否推动创新?

Evgeny Kagan, Christian Jost, Tobias Lieberum, Sebastian Schiffels

专题命中 其他LLM :prompting(abstract)

AI总结 本研究通过实验发现,迭代工作流程在创新任务中表现更优,但其优势随时间减弱,且在特定条件下效果受限。

详情

展开后加载摘要…

URL PDF HTML 收藏