arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-05-05 至 2026-05-05 共收录 29 信号源:cs.CL, cs.AI, cs.LG

1. 效率与部署 29 篇

2602.15823 2026-05-05 cs.LG cs.AI 92%

CrispEdit: Low-Curvature Projections for Scalable Non-Destructive LLM Editing

CrispEdit:可扩展的非破坏性LLM编辑中的低曲率投影

Zarif Ikram, Arad Firouzkouhi, Stephen Tu, Mahdi Soltanolkotabi, Paria Rashidinejad

机构 * University of Southern California(南加州大学)

专题命中 效率与部署 :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 CrispEdit通过低曲率子空间约束优化,实现LLM编辑中的能力保持,有效降低能力退化,提升编辑效果。

Comments ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06838 2026-05-05 cs.AR cs.LG 92%

P3-LLM: An Integrated NPU-PIM Accelerator for Edge LLM Inference Using Hybrid Numerical Formats

P3-LLM:一种用于边缘LLM推理的NPU-PIM集成加速器,采用混合数值格式

Yuzong Chen, Chao Fang, Xilai Dai, Yuheng Wu, Thierry Tambe, Marian Verhelst, Mohamed S. Abdelfattah

机构 * Department of Electrical and Computer Engineering, Cornell University(康奈尔大学电气与计算机工程系) EAST-MICAS, KU Leuven(KU莱顿大学EAST-MICAS) Department of Electrical Engineering, Stanford University(斯坦福大学电气工程系)

专题命中 效率与部署 :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.LG

AI总结 本文提出P3-LLM,通过混合精度量化和PIM架构设计,提升边缘LLM推理效率,实现4.9倍以上的加速效果。

Comments Accepted to the 53rd IEEE/ACM International Symposium on Computer Architecture (ISCA), 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00815 2026-05-05 cs.AI 91%

Resource-Efficient Reinforcement for Reasoning Large Language Models via Dynamic One-Shot Policy Refinement

通过动态单次策略细化实现大规模语言模型推理的资源高效强化学习

Yunjian Zhang, Sudong Wang, Yang Li, Peiran Xu, Conghao Zhou, Xiaoyue Ma, Jianing Li, Yao Zhu

机构 * UCAS(中国科学院大学) HKUST(GZ)(香港科技大学(广州)) Tsinghua University(清华大学) Sun Yat-Sen University(中山大学) Xidian University(西安电子科技大学) George Mason University(乔治·梅森大学) Peking University(北京大学) Zhejiang University(浙江大学)

专题命中 效率与部署 :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn);post-training(abstract)

AI总结 本文提出动态单次策略细化方法,通过减少训练样本数量和计算开销,提升大规模语言模型推理效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.02380 2026-05-05 cs.LG 91%

EntroLLM: Entropy Encoded Weight Compression for Efficient Large Language Model Inference on Edge Devices

EntroLLM:基于熵编码的权重压缩用于边缘设备上高效的大语言模型推理

Arnab Sanyal, Gourav Datta, Prithwish Mukherjee, Sandeep P. Chinchali, Michael Orshansky

机构 * The University of Texas at Austin(德克萨斯大学奥斯汀分校) Georgia Institute of Technology(佐治亚理工学院) Case Western Reserve University(凯斯西储大学)

专题命中 效率与部署 :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn);post-training(abstract)

AI总结 EntroLLM结合混合量化和熵编码技术,通过提升权重压缩率和编码效率,在边缘设备上实现高效的大语言模型推理,实验显示存储节省达30%-65%,推理速度提升31.9%-146.6%。

Comments 4 pages, 1 reference page

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.25426 2026-05-05 cs.CL cs.AI 89%

Implicature in Interaction: Understanding Implicature Improves Alignment in Human-LLM Interaction

互动中的隐含意义:理解隐含意义能提高人-大语言模型互动的对齐

Asutosh Hota, Jussi P. P. Jokinen

机构 * Jyvaskylan yliopisto(耶夫斯克扬大学)

专题命中 效率与部署 :LLM(title,summary_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 研究探讨LLM在基于上下文的提示中推断用户意图的能力,发现更大模型更接近人类解释,隐含意义提示能显著提升响应的相关性和质量,67.6%参与者偏好隐含意义提示。

Comments The manuscript is approximately 7360 words and contains 12 figures and 6 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.01373 2026-05-05 cs.CL cs.AI 88%

Focus on the Core: Empowering Diffusion Large Language Models by Self-Contrast

聚焦核心:通过自对比增强扩散大语言模型

Jinyuan Feng, Xin Yu, Yiqun Chen, Xiaochi Wei, Yan Gao, Yi Wu, Yao Hu, Zhiqiang Pu

机构 * Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) Renmin University of China(中国人民大学) Xiaohongshu Inc.(小红书公司)

专题命中 效率与部署 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文通过研究高信息密度token,提出FoCore解码策略,利用自对比提升生成质量与效率,实验表明在数学、代码和逻辑推理任务中效果显著。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.01255 2026-05-05 cs.LG 88%

Activation Compression in LLMs: Theoretical Analysis and Efficient Algorithm

LLMs中的激活压缩:理论分析与高效算法

Wen-Da Wei, Han-Bin Fang, Yang-Di Liu, Jiang-Xin Shi, James Kwok, Yu-Feng Li

机构 * Nanjing University(南京大学) Tsinghua University(清华大学) Huazhong University of Science and Technology(华中科技大学) Hong Kong University of Science and Technology(香港科学大学)

专题命中 效率与部署 :LLM(summary_cn,abstract);large language model(abstract);language model(abstract);pretraining(abstract)

AI总结 本文提出一种激活-梯度联合压缩方法,通过低秩因子压缩线性层梯度,无需额外计算,提升LLM训练效率与压缩性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.01280 2026-05-05 cs.DC cs.AI 87%

Position: LLM Serving Needs Mathematical Optimization and Algorithmic Foundations, Not Just Heuristics

位置:大语言模型服务需要数学优化和算法基础,而不是仅仅启发式方法

Zijie Zhou

机构 * Department of Industrial Engineering and Decision Analytics(工业工程与决策分析系)

专题命中 效率与部署 :LLM(title,summary_cn);分类 cs.AI

AI总结 本文指出大语言模型推理服务已超越通用启发式方法,需数学优化和算法基础。现有服务系统如vLLM和SGLang的算法核心仍沿用传统分布式计算方法,无法适应LLM推理的动态KV缓存、prefill-decode阶段不对称性等特性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00831 2026-05-05 cs.DC cs.AI cs.PF 87%

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving

GhostServe: 一种轻量级的影子中检查点系统,用于容错的大语言模型服务

Shakya Jayakody, Youpeng Zhao, Chinmay Dhanraj Nehate, Jun Wang

机构 * Anonymous Authors(匿名作者)

专题命中 效率与部署 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出GhostServe,一种轻量级检查点系统,通过擦除编码在主机内存中生成和存储奇偶分片,以保护流式KV缓存,提升大语言模型服务的容错性和效率。

Comments MLSys 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.27396 2026-05-05 cs.AR 87%

VitaLLM: A Versatile, Ultra-Compact Ternary LLM Accelerator with Dependency-Aware Scheduling

VitaLLM:一种多功能、超紧凑的三进制大语言模型加速器及其依赖感知调度

Zi-Wei Lin, Tian-Sheuan Chang

专题命中 效率与部署 :LLM(title,abstract);large language model(abstract);language model(abstract)

AI总结 本文提出VitaLLM,一种针对三进制大语言模型推理的高效加速器,通过双核计算策略和依赖感知调度框架,实现高吞吐量和低功耗。

Journal ref IEEE Transactions on Circuits and Systems for Artificial Intelligence, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.01034 2026-05-05 cs.CL 86%

A Theoretical Game of Attacks via Compositional Skills

通过组合技能的攻击理论游戏

Xinbo Wu, Huan Zhang, Abhishek Umrawal, Lav R. Varshney

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Stony Brook University(石溪大学)

专题命中 效率与部署 :LLM(summary_cn,abstract_cn);large language model(abstract);language model(abstract);prompting(abstract)

AI总结 本文提出一个理论框架,分析攻击者与防御者之间的博弈,设计最优攻击策略并揭示其与现有对抗提示方法的关系,同时推导出可证明最优的防御策略,并通过实验验证其在不同LLM和基准上的优越性。

Comments arXiv admin note: text overlap with arXiv:2505.20841

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.01278 2026-05-05 cs.AI 85%

Valley3: Scaling Omni Foundation Models for E-commerce

Valley3:面向电商的多模态大语言模型规模化

Zeyu Chen, Guanghao Zhou, Qixiang Yin, Ziwang Zhao, Huanjin Yao, Pengjiu Xia, Min Yang, Cen Chen, Minghui Qiu

机构 * Valley Team, ByteDance Group(字节跳动集团山谷团队)

专题命中 效率与部署 :foundation model(title);large language model(abstract);language model(abstract);post-training(abstract)

AI总结 Valley3通过四阶段预训练管道,融合多模态能力与电商知识,实现跨模态指令跟随和长链推理,构建了涵盖6个任务的电商基准测试,展现出在电商场景中的优越性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.01546 2026-05-05 cs.NI cs.AI 81%

6G Needs Agents: Toward Agentic AI-Native Networks for Autonomous Intelligence

6G需要智能体:迈向面向自主智能的智能体AI原生网络

Mohamed Amine Ferrag, Abderrahmane Lakas, Merouane Debbah

机构 * Department of Computer and Network Engineering, United Arab Emirates University, UAE(计算机与网络工程系,阿联酋大学) Research Institute for Digital Future, Khalifa University, UAE(数字未来研究院,哈利法大学)

专题命中 效率与部署 :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出基于大语言模型的智能体架构,用于构建AI原生6G网络,通过分层推理和分布式多智能体系统平衡性能与效率,揭示了模型异构部署和量化影响的重要性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.01416 2026-05-05 cs.CY cs.CL 81%

Who Decides What Is Harmful? Content Moderation Policy Through A Multi-Agent Personalised Inference Framework

谁决定什么是有害的?通过多智能体个性化推理框架进行内容审核政策

Ewelina Gajewska, Michal Wawer, Katarzyna Budzynska, Jaroslaw A. Chudziak

机构 * Warsaw University of Technology(华沙技术大学)

专题命中 效率与部署 :LLM(summary_cn,abstract);分类 cs.CL

AI总结 本文提出基于LLM的多智能体个性化推理框架,通过用户敏感性档案过滤内容,提升审核准确性,并为平台治理提供政策相关洞察。

Comments The paper has been accepted to the 34th European Conference on Information Systems (ECIS 2026). The official paper version will appear in the conference proceedings

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.01106 2026-05-05 cs.CL cs.AI 81%

Component-Aware Self-Speculative Decoding in Hybrid Language Models

混合语言模型中的组件感知自推测解码

Hector Borobia, Elies Seguí-Mas, Guillermina Tormo-Carbó

机构 * organization= VRAIN -- Valencian Research Institute for Artificial Intelligence, Universitat Polit\`ecnica de Val\`encia , city= Valencia , country= Spain organization= Department of Economics organization= Department of Business Organisation, Universitat Polit\`ecnica de Val\`encia , city= Valencia , country= Spain

专题命中 效率与部署 :language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文提出组件感知自推测解码,首次利用混合语言模型的内部架构异质性,通过隔离SSM/线性注意力子图作为零成本内部草案,提升了推测解码效率,并展示了不同架构混合模型的组件组合模式对自推测可行性的影响。

Comments 29 pages, 1 figure, 9 tables. Code: https://github.com/hecboar/hybrid-speculative-decoding

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00857 2026-05-05 eess.SP cs.AI cs.LG q-bio.NC 81%

Foundation Model Guided Dual-Branch Co-Adaptation for Source-Free EEG Decoding

基于基础模型的双分支共适应源无关EEG解码

Peiliang Gong, Han Zhang, Zhen Jiang, Chenyu Liu, Ziyu Jia, Xinliang Zhou, Daoqiang Zhang, Xiaoli Li

机构 * College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算与数据科学学院) College of Artifical Intelligence and Automation, Hohai University(河海大学人工智能与自动化学院) College of Artificial Intelligence, Nanjing University of Aeronautics and Astronautics(南京航空航天大学人工智能学院) Brainnetcome Center, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所脑网络中心) Information Systems Technology and Design, Singapore University of Technology and Design(新加坡科技设计大学信息系统技术与设计)

专题命中 效率与部署 :foundation model(title,abstract);分类 cs.AI、cs.LG

AI总结 本文提出FUSED框架,通过双分支共适应机制整合大规模基础模型与紧凑专家模型,提升源无关EEG解码的泛化能力和稳定性,实验验证其在多任务中的优越性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00930 2026-05-05 q-bio.GN cs.AI 79%

CellxPert: Inference-Time MCMC Steering of a Multi-Omics Single-Cell Foundation Model for In-Silico Perturbation

CellxPert:一种多组学单细胞基础模型的推理时间MCMC引导方法用于计算机模拟扰动

Andac Demir, Erik W. Anderson, Jeremy L. Jenkins, Srayanta Mukherjee

机构 * Novartis Biomedical Research(诺华生物医学研究)

专题命中 效率与部署 :foundation model(title,abstract);分类 cs.AI

AI总结 CellxPert通过整合多组学数据,实现了对单细胞和空间多组学的统一表示,支持细胞类型注释、扰动响应预测和多组学整合,采用MCMC方法提升生物可解释性。

Journal ref ICLR Machine Learning for Genomics Explorations Workshop 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.01662 2026-05-05 cs.CV 78%

Video Active Perception: Effective Inference-Time Long-Form Video Understanding with Vision-Language Models

视频主动感知:基于视觉-语言模型的高效推理时长视频理解

Martin Q. Ma, Willis Guo, Aditya Agrawal, Ankit Gupta, Paul Pu Liang, Ruslan Salakhutdinov, Louis-Philippe Morency

机构 * Carnegie Mellon University(卡内基梅隆大学) MIT(麻省理工学院)

专题命中 效率与部署 :language model(title,abstract)

AI总结 本文提出视频主动感知方法,通过主动感知理论提升长视频问答性能,实现帧效率提升5.6倍,优于现有模型。

Comments ICCV 2025 workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.01506 2026-05-05 cs.CV 75%

OmniEncoder: See, Hear, and Feel Continuous Motion Like Humans With One Encoder

OmniEncoder: 通过一个编码器实现如同人类般连续的视觉、听觉与触觉感知

Detao Bai, Shimin Yao, Weixuan Chen, Chengen Lai, Yuanming Li, Zhiheng Ma, Xihan Wei

机构 * Tongyi Lab Alibaba Group(阿里巴巴集团通义实验室) Shenzhen University of Advanced Technology(深圳先进技术大学)

专题命中 效率与部署 :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 OmniEncoder通过统一的Transformer架构,在共享潜在空间中对视觉和音频信号以25fps对称嵌入,解决多模态解耦与计算效率问题,提升连续视觉理解任务性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.01441 2026-05-05 cs.CL cs.CY cs.HC 70%

Artificial intelligence language technologies in multilingual healthcare: Grand challenges ahead

多语言医疗中的人工智能语言技术:前方的宏大挑战

Vicent Briva-Iglesias

机构 * School of Applied Languages and Intercultural Studies (SALIS)(应用语言学与跨文化研究学院) CTTS, ADAPT Centre(CTTS与ADAPT中心) Dublin City University(都柏林城市大学)

专题命中 效率与部署 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文探讨多语言医疗中AI语言技术的应用挑战,分析其在翻译、文档等任务中的表现差异及安全性和公平性问题,提出七大研究与部署挑战。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.01078 2026-05-05 cs.CR cs.AI 70%

A Sentence Relation-Based Approach to Sanitizing Malicious Instructions

基于句子关系的方法用于清除恶意指令

Soumil Datta, Melissa Umble, Daniel S. Brown, Guanhong Tao

机构 * University of Utah(犹他大学)

专题命中 效率与部署 :LLM(abstract,abstract_cn);分类 cs.AI

AI总结 本文提出SONAR框架,通过自然语言推理指标识别并移除恶意指令,有效降低攻击成功率,优于现有基线防御方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00955 2026-05-05 cs.CR cs.AI 70%

E-MIA: Exam-Style Black-Box Membership Inference Attacks against RAG Systems

E-MIA:针对RAG系统的考试风格黑盒成员推断攻击

Zelin Guan, Shengda Zhuo, Zeyan Li, Jinchun He, Wangjie Qiu, Zhiming Zheng, Shuqiang Huang

机构 * College of Cyber Security, Jinan University(济南大学网络安全学院) School of Computer Science, Shanghai Jiaotong University(上海交通大学计算机科学学院) Institute of Artificial Intelligence, Beijing Advanced Innovation Center for Future Blockchain and Privacy Computing, Beihang University(北京航空航天大学人工智能研究院) Zhongguancun Laboratory, Beijing(中关村实验室)

专题命中 效率与部署 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出E-MIA,通过将目标文档中的可验证硬证据转化为包含四种客观评分题型的考试,利用多证据目标问题的综合考试分数作为成员信号,提升在严格设置下的成员/非成员分离能力,同时保持自然隐蔽的查询。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17616 2026-05-05 cs.LG 70%

Split-on-Share: Mixture of Sparse Experts for Task-Agnostic Continual Learning

Split-on-Share:任务无关持续学习的稀疏专家混合

Fatema Siddika, Md Anwar Hossen, Tanwi Mallick, Ali Jannesari

专题命中 效率与部署 :large language model(abstract);language model(abstract);分类 cs.LG

AI总结 SETA通过分解模型为模块化子空间解决持续学习中的可塑性-稳定性矛盾,通过弹性权重锚定保护共享知识并自动检索任务专用专家组合,优于现有参数高效微调方法。

Comments we are updating the paper and will release another version soon

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.01644 2026-05-05 cs.CR 67%

Toward a Principled Framework for Agent Safety Measurement

迈向代理安全测量的原理性框架

Shuyi Lin, Anshuman Suri, Alina Oprea, Cheng Tan

专题命中 效率与部署 :LLM(abstract,abstract_cn)

AI总结 本文提出基于搜索而非采样的原理性框架,用于评估代理安全,通过BOA框架在预算内搜索轨迹空间,发现贪心和采样方法遗漏的不安全轨迹。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00943 2026-05-05 cs.RO 67%

ARIS: Agentic and Relationship Intelligence System for Social Robots

ARIS:面向社交机器人的代理与关系智能系统

Stavya Datta, Fucai Ke, Leimin Tian, Hamid Rezatofighi

机构 * Monash University(墨尔本大学)

专题命中 效率与部署 :large language model(abstract);language model(abstract)

AI总结 ARIS通过整合多模态推理、图基社会世界模型和检索增强生成技术,提升社交机器人在多轮交互和社会关系推理中的能力,实验显示其在智能感知和用户亲和力方面表现更优。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00839 2026-05-05 cs.AI cs.LG 62%

2026 Roadmap on Artificial Intelligence and Machine Learning for Smart Manufacturing

2026 年人工智能与机器学习在智能制造中的路线图

Jay Lee, Hanqi Su, Marco Macchi, Adalberto Polenghi, Wei Wu, Zhiheng Zhao, George Q. Huang, Kiva Allgood, Devendra Jain, Benedikt Gieger, Vibhor Pandhare, Soumyabrata Bhattacharjee, Ram Mohril, Lingbao Kong, Qiyuan Wang, Xinlan Tang, Sungjong Kim, Chan Hee Park, Byeng D. Youn, Guo Dong Goh, Xi Huang, Wai Yee Yeong, Yung C Shin, He Zhang, Zitong Wang, Fei Tao, Jagjit Singh Srai, Satyandra K. Gupta, Byung Gun Joung, Albin John, John W. Sutherland, Sang Won Lee, Olga Fink, Vinay Sharma, Faez Ahmed, Wei Chen, Mark Fuge, Arild Waaler, Martin G. Skjæveland, Dimitris Kyritsis, Wei Chen, VispiNevile Karkaria, Yi-Ping Chen, Ying-Kuan Tsai, Joseph Cohen, Xun Huan, Jing Lin, Liangwei Zhang, Gregory W. Vogl, Aaron W. Cornelius, Xiaodong Jia, Dai-Yan Ji, Takanobu Minami, Ruoxin Wang

机构 * Center for Industrial Artificial Intelligence, Department of Mechanical Engineering, University of Maryland, College Park(工业人工智能中心,机械工程系,马里兰大学College Park分校) Department of Management, Economics and Industrial Engineering, Politecnico di Milano(管理、经济与工业工程系,米兰理工学院) Department of Industrial and Systems Engineering, The Hong Kong Polytechnic University(工业与系统工程系,香港理工大学) Centre for Advanced Manufacturing & Supply Chains, World Economic Forum(先进制造与供应链研究中心,世界经济论坛) Department of Mechanical Engineering, Indian Institute of Technology Indore(机械工程系,印度理工学院Indore分校) Future Information Innovative College, Fudan University(未来信息创新学院,复旦大学) Department of Mechanical Engineering, Seoul National University(机械工程系,首尔国立大学) Department of Mechanical and Information Engineering, University of Seoul(机械与信息工程系,首尔大学) Onepredict Corp.(Onepredict公司) School of Mechanical and Aerospace Engineering, Nanyang Technological University(机械与航空航天工程学院,南洋理工大学) Singapore Centre for 3D Printing, Nanyang Technological University(新加坡3D打印中心,南洋理工大学) Mechanical Engineering, Purdue University(机械工程系,普渡大学) Digital Twin International Research Center, International Institute for Interdisciplinary and Frontiers, Beihang University(数字孪生国际研究中心, interdisciplinary and Frontiers 国际研究院,北京航空航天大学) School of Automation Science and Electrical Engineering, Beihang University(自动化科学与电气工程学院,北京航空航天大学) Department of Engineering, University of Cambridge(工程系,剑桥大学) Center for Advanced Manufacturing, University of Southern California(先进制造中心,南加州大学) School of Sustainability Engineering and Environmental Engineering, Purdue University(可持续工程与环境工程系,普渡大学) School of Mechanical Engineering, Sungkyunkwan University(机械工程系,全南大学) Intelligent Maintenance and Operations Systems, EPFL(智能维护与运营系统,苏黎世联邦理工学院) Department of Mechanical Engineering, Massachusetts Institute of Technology(机械工程系,麻省理工学院) J. Mike Walker ’66 Department of Mechanical Engineering, Texas A&M University(J. Mike Walker ’66 机械工程系,德克萨斯A&M大学) Department of Mechanical and Process Engineering, ETH Zürich(机械与工艺工程系,苏黎世联邦理工学院)

专题命中 效率与部署 :foundation model(abstract);分类 cs.AI、cs.LG

AI总结 本文探讨人工智能与机器学习在智能制造中的发展现状与未来方向,涵盖基础理论、应用领域及新兴技术,旨在推动创新与产业应用。

Comments This paper has been accepted for publication in the Journal Machine Learning: Engineering

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.01325 2026-05-05 cs.CV cs.LG 57%

Rethinking Model Selection in VLM Through the Lens of Gromov-Wasserstein Distance

通过格罗莫夫-瓦瑟斯坦距离重新思考VLM中的模型选择

Muyang Li, Yucheng Liu, Jianbo Ma, Elliot Osborne, Bo Han, Tongliang Liu

机构 * Sydney AI Centre, The University of Sydney(悉尼人工智能中心,悉尼大学) Dolby Laboratories(杜比实验室) TMLR Group, Hong Kong Baptist University(香港 Baptist 大学 TMLR 团体)

专题命中 效率与部署 :language model(abstract);分类 cs.LG

AI总结 本文通过系统实验探讨视觉编码器选择的关键因素,发现模态间结构相似性通过格罗莫夫-瓦瑟斯坦距离衡量能提升VLM性能预测。

Comments Accepted as Highlight publication for CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.05254 2026-05-05 cs.IR cs.CL 57%

TagRAG: Tag-guided Hierarchical Knowledge Graph Retrieval-Augmented Generation

TagRAG:基于标签的分层知识图谱检索增强生成

Wenbiao Tao, Xinyuan Li, Yunshi Lan, Weining Qian

机构 * East China Normal University(东华师范大学)

专题命中 效率与部署 :language model(abstract);分类 cs.CL

AI总结 TagRAG通过构建标签知识图谱和标签引导的检索生成框架,提升小语言模型的全局推理能力和知识增量效率,实验显示其在多个领域数据集上表现优异。

Comments Accepted by ACL 2026 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.01330 2026-05-05 cs.CV 50%

Colinearity Decay: Training Quantization-Friendly ViTs with Outlier Decay

共线性衰减:通过异常衰减训练量化友好的视觉Transformer

Jin Tong, Guang Liang, Peilin Sun, Jianxin Wu

机构 * State Key Laboratory of Novel Software Technology, Nanjing University(南京大学新型软件技术国家重点实验室) School of Artificial Intelligence, Nanjing University(南京大学人工智能学院) Zhongguancun Academy, Beijing(中关村学院)

专题命中 效率与部署 :post-training(abstract)

AI总结 本文提出Colinearity-Decay方法,通过结构正则化控制Transformer块内的矩阵对对齐,减少极端激活,提升低比特部署精度,同时保持全精度性能。

Comments 17 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏