arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12169 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12169 篇

2503.10566 2026-02-10 cs.LG 83%

ASIDE: Architectural Separation of Instructions and Data in Language Models

ASIDE: 语言模型中指令与数据的架构分离

Egor Zverev, Evgenii Kortukov, Alexander Panfilov, Alexandra Volkova, Soroush Tabesh, Sebastian Lapuschkin, Wojciech Samek, Christoph H. Lampert

机构 * Institute of Science and Technology Austria (ISTA)(奥地利科学与技术研究所) Fraunhofer Heinrich Hertz Institute(弗劳恩霍夫海因里希·赫兹研究所) ELLIS Institute Tübingen(图宾根ELLIS研究所) Max Planck Institute for Intelligent Systems(马克斯·普朗克智能系统研究所) Tübingen AI Center(图宾根人工智能中心) Centre of eXplainable Artificial Intelligence(可解释人工智能中心) Technische Universität Berlin(柏林技术大学) Berlin Institute for the Foundations of Learning and Data (BIFOLD)(柏林学习与数据基础研究所)

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.LG

AI总结 ASIDE通过在令牌嵌入层面实现指令与数据的分离,提升了语言模型的安全性和性能

Comments ICLR 2026 paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18727 2026-02-10 cs.LG 83%

LogSyn: A Few-Shot LLM Framework for Structured Insight Extraction from Unstructured General Aviation Maintenance Logs

LogSyn: 一种基于少样本学习的LLM框架,用于从非结构化通用航空维护日志中提取结构化洞察

Devansh Agarwal, Maitreyi Chatterjee, Biplab Chatterjee

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.LG

AI总结 LogSyn通过少样本学习利用LLM将非结构化航空维护日志转化为结构化数据,实现故障模式识别与事件分类,提升航空维护流程和预测分析能力。

Comments Accepted in Proceedings of the 3rd INCOM 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.07053 2026-02-06 cs.CL cs.SD eess.AS 83%

TASTE: Text-Aligned Speech Tokenization and Embedding for Spoken Language Modeling

TASTE: 用于语音语言建模的文本对齐语音标记化与嵌入

Liang-Hsuan Tseng, Yi-Chang Chen, Kuan-Yi Lee, Da-Shan Shiu, Hung-yi Lee

机构 * MediaTek Research(联发科技研究) Graduate Institute of Communication Engineering, National Taiwan University(国立台湾大学通信工程研究所) Artificial Intelligence Center of Research Excellence, National Taiwan University(国立台湾大学卓越人工智能研究中心)

专题命中 其他LLM :language model(title,abstract);LLM(abstract);分类 cs.CL

AI总结 TASTE通过语音重建目标实现文本对齐的语音标记化与嵌入,提升语音语言建模的性能。

Comments ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.22585 2026-02-02 cs.DC cs.LG 83%

HetCCL: Accelerating LLM Training with Heterogeneous GPUs

HetCCL: 通过异构GPU加速大语言模型训练

Heehoon Kim, Jaehwan Lee, Taejeoung Kim, Jongwon Park, Jinpyo Kim, Pyongwon Suh, Ryan H. Choi, Sangwoo Lee, Jaejin Lee

机构 * snu-cse(首尔国立大学计算机科学与工程系) snu-ds(首尔国立大学数据科学研究院) moreh(Moreh公司) samsung(三星研究)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.LG

AI总结 HetCCL通过统一异构GPU的集体通信库,实现跨供应商的高效训练,提升异构环境下的大语言模型训练性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.14803 2026-01-08 cs.CY cs.AI 83%

OnlineMate: An LLM-Based Multi-Agent Companion System for Cognitive Support in Online Learning

OnlineMate: 基于大语言模型的多智能体伴侣系统用于在线学习中的认知支持

Xian Gao, Zongyun Zhang, Ting Liu, Yuzhuo Fu

机构 * Shanghai Jiao Tong University(上海交通大学)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 OnlineMate通过基于LLM的多智能体系统,结合理论思维,为在线学习提供个性化认知支持,提升学习深度与情感参与。

Comments work in progress

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.14435 2026-01-06 cs.CY cs.AI 83%

Choosing a Model, Shaping a Future: Comparing LLM Perspectives on Sustainability and its Relationship with AI

选择模型,塑造未来:比较LLM对可持续性和其与AI关系的视角

Annika Bush, Meltem Aksoy, Markus Pauly, Greta Ontrup

机构 * Research Center Trustworthy Data Science and Security, University Alliance Ruhr(可信数据科学与安全研究中心,鲁尔大学联盟) Department of Computer Science, Technical University Dortmund(计算机科学系,图林根技术大学) Chair of Mathematical Statistics and Applications in Industry, Technical University Dortmund(工业数学统计与应用教授职位,图林根技术大学) Department of Computer Science, University of Duisburg-Essen(计算机科学系,杜伊斯堡-埃森大学)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本研究比较了五个先进LLM对可持续性与AI关系的视角,发现模型间存在显著差异,强调模型选择对可持续性战略的影响。

Comments Accepted for EMNLP Conference

Journal ref Conference on Empirical Methods in Natural Language Processing (EMNLP 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.10904 2025-12-29 cs.CR cs.AI 83%

CEKER: A Generalizable LLM Framework for Literature Analysis with a Case Study in Unikernel Security

CEKER:一种通用的LLM框架用于文献分析及在unikernel安全领域的案例研究

Alex Wollman, John Hastings

机构 * The Beacom College of Computer and Cyber Sciences(计算机与网络安全科学学院) Dakota State University(达科他州立大学)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 CEKER通过LLM框架实现文献分析自动化,应用于unikernel安全领域,揭示了攻击面减少及安全缺口,强调动态安全调整的重要性。

Comments 7 pages, 2 figures

Journal ref International Symposium on Intelligent Computing and Networking 2025 (ISICN 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.19950 2025-12-24 cs.CL cs.HC 83%

Bias Beneath the Tone: Empirical Characterisation of Tone Bias in LLM-Driven UX Systems

音调下的偏见:LLM驱动的UX系统中音调偏见的实证刻画

Heet Bodara, Md Masum Mushfiq, Isma Farah Siddiqui

机构 * Faculty of Information Technology, Monash University(信息技术学院,莫纳什大学)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文研究了LLM驱动UX系统中的音调偏见问题,通过合成数据集和弱监督方法,揭示了模型在对话中存在系统性的音调偏差,为设计公平可信的对话AI提供了依据。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12706 2025-12-16 cs.AI cs.SE 83%

Synergizing Code Coverage and Gameplay Intent: Coverage-Aware Game Playtesting with LLM-Guided Reinforcement Learning

协同代码覆盖率与游戏意图:基于LLM引导强化学习的覆盖感知游戏测试

Enhong Mu, Minami Yoda, Yan Zhang, Mingyue Zhang, Yutaka Matsuno, Jialong Li

机构 * College of Computer and Information Science, Southwest University(西南大学计算机与信息科学学院) College of Science and Technology, Nihon University(日本立命馆大学科学技术学院) Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院) Waseda Institute for Advanced Study, Waseda University(早稻田大学高级研究机构)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 SMART框架通过LLM引导强化学习,结合结构验证与功能验证,提升游戏更新测试的覆盖率和任务完成率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12630 2025-12-16 cs.HC cs.AI 83%

ORIBA: Exploring LLM-Driven Role-Play Chatbot as a Creativity Support Tool for Original Character Artists

ORIBA:探索基于大语言模型的对话机器人作为原创角色艺术家创造力支持工具

Yuqian Sun, Xingyu Li, Shunyu Yao, Noura Howell, Tristan Braud, Chang Hee Lee, Ali Asadipour

机构 * Computer Science Research Centre, Royal College of Art(皇家艺术学院计算机科学研究中心) Digital Media, School of Literature, Media, and Communication, Georgia Institute of Technology(佐治亚理工学院数字媒体系) Princeton University(普林斯顿大学) Digital Media, Georgia Institute of Technology(佐治亚理工学院数字媒体系) Division of Integrative Systems and Design, The Hong Kong University of Science and Technology(香港科学大学整合系统与设计 division) Industrial Design Department, College of Engineering, KAIST(韩国科学技术院工程学院工业设计系)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 ORIBA通过大语言模型支持原创角色艺术家的创意过程,平衡AI辅助与创意自主权。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08951 2025-12-11 cs.NE cs.AI cs.GR 83%

AI Co-Artist: A LLM-Powered Framework for Interactive GLSL Shader Animation Evolution

AI共艺术家:一种基于大语言模型的交互式GLSL着色器动画进化的框架

Kamer Ali Yuksel, Hassan Sawaf

机构 * aiXplain Inc.(aiXplain公司)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 AI Co-Artist利用大语言模型降低GLSL着色器创作门槛,通过直观交互实现视觉艺术的进化与生成。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07312 2025-12-09 cs.AR cs.AI cs.DC 83%

DCO: Dynamic Cache Orchestration for LLM Accelerators through Predictive Management

DCO: 通过预测管理实现LLM加速器的动态缓存编排

Zhongchun Zhou, Chengtao Lai, Yuhang Gu, Wei Zhang

机构 * Department of Electronic and Computer Engineering, The Hong Kong University of Science and Technology(电子与计算机工程系,香港科学与技术大学) School of Electronic Science and Engineering, Southeast University(电子科学与工程学院,东南大学)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 DCO通过预测管理实现LLM加速器的动态缓存编排,利用数据流信息优化缓存替换和旁路决策,提升性能至1.8倍,面积仅0.064mm²。

Comments \c{opyright} 2025 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work in other works

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14070 2025-11-27 cs.CR cs.AI 83%

Special-Character Adversarial Attacks on Open-Source Language Model

特殊字符对抗攻击对开源语言模型的攻击

Ephraiem Sarabamoun

机构 * University of Virginia(弗吉尼亚大学)

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.AI

AI总结 本文研究了针对开源语言模型的特殊字符对抗攻击,评估了多种攻击方式并揭示了模型的安全漏洞及失败模式。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.20627 2025-11-26 cs.AI 83%

Fighting AI with AI: Leveraging Foundation Models for Assuring AI-Enabled Safety-Critical Systems

用AI对抗AI:利用基础模型确保AI赋能的安全关键系统

Anastasia Mavridou, Divya Gopinath, Corina S. Păsăreanu

机构 * KBR Inc.(KBR公司) NASA Ames(美国国家航空航天局阿姆斯研究中心)

专题命中 其他LLM :foundation model(title);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出利用AI技术解决安全关键系统中AI保证问题,通过REACT和SemaLens两个组件实现需求工程与感知系统验证。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18194 2025-11-25 cs.CL 83%

Agent-as-a-Graph: Knowledge Graph-Based Tool and Agent Retrieval for LLM Multi-Agent Systems

Agent-as-a-Graph: 基于知识图谱的LLM多智能体系统中的工具与智能体检索

Faheem Nizar, Elias Lumer, Anmol Gulati, Pradeep Honaganahalli Basavaraju, Vamse Kumar Subbiah

机构 * Commercial Technology and Innovation Office, PricewaterhouseCoopers(普华永道商业技术与创新办公室)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出基于知识图谱的智能体检索方法,通过构建智能体与工具的关系图谱,提升多智能体系统中工具和智能体的检索效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.15478 2025-11-25 cs.CL 83%

Red Teaming Multimodal Language Models: Evaluating Harm Across Prompt Modalities and Models

针对多模态语言模型的红队测试:评估不同提示模态和模型的有害性

Madison Van Doren, Casey Ford

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.CL

AI总结 本研究通过红队测试评估多模态语言模型在不同提示模态下的安全性,发现Pixtral 12B的有害响应率最高,而Claude Sonnet 3.5最安全,凸显了建立多模态安全基准的必要性。

Journal ref AAAI 2026 AIGOV Workshop and EurIPS 2025 Workshop on Unifying Perspectives on Learning Biases

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08535 2025-11-12 cs.CV cs.AI 83%

Large Sign Language Models: Toward 3D American Sign Language Translation

Sen Zhang, Xiaoxiao He, Di Liu, Zhaoyang Xia, Mingyu Zhao, Chaowei Tan, Vivian Li, Bo Liu, Dimitris N. Metaxas, Mubbasir Kapadia

机构 * Rutgers University(罗格斯大学) Meta Reality Labs(Meta现实实验室) Qualcomm(高通公司) Walmart Global Tech(沃尔玛全球技术) Roblox PRISMS

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00854 2025-11-04 cs.CL 83%

TriCon-Fair: Triplet Contrastive Learning for Mitigating Social Bias in Pre-trained Language Models

Chong Lyu, Lin Li, Shiqing Wu, Jingling Yuan

机构 * School of Computer Science(计算机科学学院) Artificial Intelligence, Wuhan University of Technology, Wuhan, China(武汉理工大学人工智能学院) Faculty of Data Science, City University of Macau, Macau, China(澳门城市大学数据科学学院)

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27016 2025-11-03 cs.CL 83%

Semantically-Aware LLM Agent to Enhance Privacy in Conversational AI Services

Jayden Serenari, Stephen Lee

机构 * Department of Computer Science University of Pittsburgh Pittsburgh, Pennsylvania, USA(计算机科学系 纽约大学 伯利恒,宾夕法尼亚州,美国)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted to IEEE Big Data 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.23271 2025-10-28 cs.CL 83%

Mubeen AI: A Specialized Arabic Language Model for Heritage Preservation and User Intent Understanding

Mohammed Aljafari, Ismail Alturki, Ahmed Mori, Yehya Kadumi

专题命中 其他LLM :language model(title,abstract);prompting(abstract);分类 cs.CL

Comments 21 pages, 2 figures, 3 tables. Includes appendices on ethical guidelines and training framework. Submitted September 04, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.08222 2025-10-28 cs.CV cs.AI cs.CY cs.HC 83%

Refusal as Silence: Gendered Disparities in Vision-Language Model Responses

Sha Luo, Sang Jung Kim, Zening Duan, Kaiping Chen

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.04510 2025-10-24 cs.CL 83%

Heterogeneous Swarms: Jointly Optimizing Model Roles and Weights for Multi-LLM Systems

Shangbin Feng, Zifeng Wang, Palash Goyal, Yike Wang, Weijia Shi, Huang Xia, Hamid Palangi, Luke Zettlemoyer, Yulia Tsvetkov, Chen-Yu Lee, Tomas Pfister

机构 * University of Washington(华盛顿大学) Google Cloud AI Research(谷歌云人工智能研究)

专题命中 其他LLM :LLM(title,abstract);language model(abstract);分类 cs.CL

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.08164 2025-10-21 cs.LG 83%

BLUR: A Bi-Level Optimization Approach for LLM Unlearning

Hadi Reisizadeh, Jinghan Jia, Zhiqi Bu, Bhanukiran Vinzamuri, Anil Ramakrishna, Kai-Wei Chang, Volkan Cevher, Sijia Liu, Mingyi Hong

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.06911 2025-10-09 cs.AI 83%

LLM-Assisted Modeling of Semantic Web-Enabled Multi-Agents Systems with AJAN

Hacane Hechehouche, Andre Antakli, Matthias Klusch

机构 * Daimler Buses GmbH(戴姆勒巴士公司) German Research Center for Artificial Intelligence (DFKI)(德国人工智能研究中心)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25286 2025-10-07 cs.CY cs.AI 83%

Artificial Authority: From Machine Minds to Political Alignments. An Experimental Analysis of Democratic and Autocratic Biases in Large-Language Models

Natalia Ożegalska-Łukasik, Szymon Łukasik

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23158 2025-10-07 cs.CL 83%

User Feedback in Human-LLM Dialogues: A Lens to Understand Users But Noisy as a Learning Signal

Yuhan Liu, Michael J. Q. Zhang, Eunsol Choi

机构 * New York University(纽约大学)

专题命中 其他LLM :LLM(title,abstract);language model(abstract);分类 cs.CL

Comments EMNLP camera-ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.00714 2025-10-07 cs.SE cs.AI 83%

RFCAudit: An LLM Agent for Functional Bug Detection in Network Protocols

Mingwei Zheng, Chengpeng Wang, Xuwei Liu, Jinyao Guo, Shiwei Feng, Xiangyu Zhang

机构 * Department of Computer Science, Purdue University(计算机科学系,普渡大学)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01929 2025-10-03 cs.CL 83%

Inverse Language Modeling towards Robust and Grounded LLMs

Davide Gabrielli, Simone Sestito, Iacopo Masi

机构 * Sapienza University of Rome(罗马萨皮恩扎大学) OmnAI Lab(OmnAI实验室)

专题命中 其他LLM :language model(title,abstract);LLM(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00047 2025-10-02 cs.CV cs.AI 83%

Explanation-Driven Counterfactual Testing for Faithfulness in Vision-Language Model Explanations

Sihao Ding, Santosh Vasa, Aditi Ramadwar

机构 * Mercedes-Benz Research & Development North America(梅赛德斯-奔驰研究与开发北美)

专题命中 其他LLM :language model(title,abstract);LLM(abstract);分类 cs.AI

Comments NeurIPS 2025 workshop on Regulatable ML

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26181 2025-10-02 cs.CL 83%

Explaining novel senses using definition generation with open language models

Mariia Fedorova, Andrey Kutuzov, Francesco Periti, Yves Scherrer

机构 * University of Oslo(奥斯陆大学) KU Leuven - Flanders Make(库尔勒文大学-佛兰德斯制造)

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.CL

Comments EMNLP 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏