arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

Northeastern University(东北大学)

共收录 1238
2602.18767 2026-07-01 cs.LO cs.LG 版本更新

Nazrin: An Atomic Neural Proof Automation Tactic in Lean 4

Nazrin: Lean 4中的原子神经证明自动化策略

Leni Aniva, Iori Oikawa, David Dill, Clark Barrett

机构 * Stanford University(斯坦福大学) Northeastern University(东北大学)

AI总结 提出原子策略集、转置原子化算法、ExprGraph数据结构和基于图神经网络的Nazrin证明器,通过仅调度原子策略克服现有证明代理的挑战,并在消费级硬件上训练和评估。

Comments 16 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.30492 2026-06-30 cs.CV

RBE-Flow: Recurrent Bayesian Estimation on Feature Manifolds for Cross-Modal Registration

RBE-Flow:基于特征流形的递归贝叶斯估计用于跨模态配准

Mengzhu Ding, Xin Song, Xiaoke Ding, Hongwei Ding, Xuecong Liu

机构 * Northeastern University, China(东北大学)

AI总结 针对跨模态图像配准中非线性辐射差异和几何畸变导致的优化困难,提出RBE-Flow框架,将密集流估计重构为特征流形上的闭环递归贝叶斯估计,通过递归流形优化和不确定性自适应概率更新实现自校正,在多个基准上取得最优性能。

Comments Accepted to ECCV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.29579 2026-06-30 cs.CV cs.AI cs.LG cs.MM

ScAle: Attention Head Scaling as a Minimal Adapter for Spatial Reasoning in Vision Language Models

ScAle: 注意力头缩放作为视觉语言模型中空间推理的最小适配器

Rahul Chowdhury, Timothy A Rupprecht, Xuan Shen, Pu Zhao, Yanzhi Wang

机构 * Northeastern University(东北大学) EmbodyX Inc.(EmbodyX公司) Zhejiang University(浙江大学)

AI总结 提出ScAle,一种超轻量适配方法,通过学习少量标量系数调节冻结骨干网络中的注意力头和MLP激活,仅用1K可训练参数在空间推理任务上实现高达134.1%的相对精度提升。

Comments Accepted by ECCV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.14252 2026-06-30 cs.LG cs.AI

Not All Timesteps Matter Equally: Selective Alignment Knowledge Distillation for Spiking Neural Networks

并非所有时间步都同等重要:用于脉冲神经网络的选择性对齐知识蒸馏

Kai Sun, Peibo Duan, Yongsheng Huang, Guowei Zhang, Benjamin Smith, Nanxu Gong, Levin Kuhlmann

机构 * Faculty of Information Technolody, Monash University, Australia(墨尔本大学信息科技学院,澳大利亚) School of Software, Northeastern University, China(东北大学软件学院,中国) Department of Medicine, National University of Singapore, Singapore(新加坡国立大学医学部,新加坡)

AI总结 本文提出Selective Alignment Knowledge Distillation方法,通过选择性对齐类别和时间知识,改进SNN性能,实验证明在静态图像和神经形态事件数据集上优于现有方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.05689 2026-06-30 cs.CV cs.AI

CRFT: Consistent-Recurrent Feature Flow Transformer for Cross-Modal Image Registration

CRFT:用于跨模态图像配准的一致-递归特征流变换器

Xuecong Liu, Mengzhu Ding, Zixuan Sun, Zhang Li, Xichao Teng

机构 * Northeastern University, China(东北大学(中国)) National University of Defense Technology, China(国防科技大学(中国))

AI总结 CRFT提出了一种基于特征流学习的统一粗到细框架,通过特征对齐和流估计实现鲁棒的跨模态图像配准,通过多尺度特征相关性和层次特征融合提升精度和鲁棒性。

Comments Accepted to CVPR 2026

Journal ref Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2026, pp. 34784-34794

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.27786 2026-06-29 cs.CL cs.AI 新提交

SHIFT: Gate-Modulated Activation Steering for Knowledge Conflict Mitigation in Retrieval-Augmented Generation

SHIFT: 门控调制激活引导用于检索增强生成中的知识冲突缓解

Ruochang Li, Pengcheng Huang, Zhenghao Liu, Yukun Yan, Huiyuan Xie, Yu Gu, Ge Yu, Maosong Sun

机构 * School of Computer Science and Engineering, Northeastern University, Shenyang, China(东北大学计算机科学与工程学院,沈阳,中国) Department of Computer Science and Technology, Tsinghua University, Beijing, China(清华大学计算机科学与技术系,北京,中国)

AI总结 提出SHIFT框架,通过轻量级门控模块调节LLM内部激活,以自适应解决检索上下文与参数知识间的冲突,仅优化0.01%参数,在六个数据集上验证有效性。

Comments 19 pages, 13 Figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.26432 2026-06-26 cs.LG econ.EM 新提交

Embedding Foundation Model Predictions in Discrete-Choice Models with Structural Guarantees

具有结构保证的嵌入基础模型预测的离散选择模型

Yingshuo Wang, Xian Sun, Yanhang Li, Zhichao Fan, Zexin Zhuang

机构 * University of California, Berkeley(加州大学伯克利分校) Duke University(杜克大学) Northeastern University(东北大学) University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Southern Methodist University(南卫理公会大学)

AI总结 提出两阶段适配器,将基础模型的预测嵌入多项Logit效用中,通过符号约束保持边际替代率,在三个数据集上平均提升6.4个百分点准确率,并保证成本单调性和合理时间价值。

Comments Extends arXiv:2605.26559 (ICML 2026 FMSD Workshop)

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.25866 2026-06-26 cs.CL 版本更新

Why Are Some Emotions Harder for LLMs? Uncovering the Causal Mechanisms of Emotion Inference via Sparse Autoencoders

从语法到情感:对LLM中情感推理的机制分析

Bangzhao Shu, Arinjay Singh, Mai ElSherief

机构 * Northeastern University(东北大学)

AI总结 本文通过稀疏自编码器研究LLM内部情感识别机制,发现情感特征在最终阶段出现,包含共享和特定情感特征,并提出可解释的因果特征引导方法提升情感识别性能。

Comments 19 pages including appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.13632 2026-06-26 cs.AI

Resilient Routing: Risk-Aware Dynamic Routing in Smart Logistics via Spatiotemporal Graph Learning

鲁棒路由:通过时空图学习实现智能物流中的风险感知动态路由

Zhiming Xue, Sichen Zhao, Yalun Qi, Xianling Zeng, Zihan Yu

机构 * College of Engineering, Northeastern University, Boston, USA(东北大学工程学院,美国波士顿) Khoury College of Computer Sciences, Northeastern University, Boston, USA(东北大学计算机科学学院,美国波士顿) College of Professional Studies, Northeastern University, Boston, USA(东北大学专业研究学院,美国波士顿)

AI总结 本文提出一种整合时空图神经网络与组合优化的鲁棒动态路由框架,通过空间聚类方法构建物流拓扑图,并利用图卷积网络和门控循环单元提取空间相关性和时间依赖性,预测拥堵风险以优化路径规划。

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.21221 2026-06-26 cs.RO 版本更新

Uncertainty-Aware Ankle Exoskeleton Control

不确定性感知的踝关节外骨骼控制

Fatima Mumtaza Tourk, Bishoy Galoaa, Sanat Shajan, Aaron J. Young, Michael Everett, Max K. Shepherd

机构 * Department of Mechanical Engineering, Northeastern University(东北大学机械工程系) College of Engineering, Northeastern University(东北大学工程学院) Khoury College of Computer Sciences, Northeastern University(东北大学计算机科学学院) Department of Electrical & Computer Engineering, Khoury College of Computer Sciences, and the Institute for Experiential Robotics, Northeastern University(计算机科学学院电子与计算机工程系及体验机器人研究所,东北大学) Woodruff School of Mechanical Engineering and the Institute for Robotics and Intelligent Machines, Georgia Institute of Technology(沃德夫机械工程学院及机器人与智能机器研究所,佐治亚理工学院) Department of Mechanical Engineering, the Department of Physical Therapy, Movement, and Rehabilitation Science, and the Institute for Experiential Robotics, Northeastern University(机械工程系、物理治疗、运动与康复科学系及体验机器人研究所,东北大学)

AI总结 提出一种不确定性感知控制框架,通过不确定性估计器自动识别不熟悉动作并关闭辅助,使踝关节外骨骼能在多种场景下安全运行,在线测试F1分数达89.2。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.25317 2026-06-25 cs.CV cs.AI 新提交

ESTANet: Efficient Online Error Detection in Procedural Videos via Prediction Inconsistency

ESTANet: 通过预测不一致性实现程序性视频中的高效在线错误检测

Shih-Po Lee, Reza Ghoddoosian, Faizan Siddiqui, Enna Sachdeva, Behzad Dariush

机构 * Honda Research Institute, USA(本田研究所(美国)) Northeastern University(东北大学)

AI总结 提出ESTANet轻量框架,利用多个动作检测器在正确与错误执行时的预测不一致性,通过多数投票实现实时在线错误检测,在三个数据集上达到最优性能。

Comments 18 pages, 8 figures, uses eccv.sty

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.24679 2026-06-25 cs.LG cs.AI 新提交

FlowPipe: LLM-Enhanced Conditional Generative Flow Networks for Data Preparation Pipeline Construction

FlowPipe: 基于条件生成流网络的LLM增强数据准备流水线构建

Kunyu Ni, Lei Cao, Jie He, Xiaotong Zhang, Jianfeng Jin, Junyu Dong, Yanwei Yu

机构 * Ocean University of China(中国海洋大学) University of Arizona(亚利桑那大学) University of Science and Technology Beijing(北京科技大学) Northeastern University(东北大学)

AI总结 提出FlowPipe框架,利用条件生成流网络(C-GFlowNets)和LLM语义调制,解决数据准备流水线自动构建中的组合优化和稀疏搜索问题,在74个数据集上平均准确率提升11.96%,训练收敛速度提升12.5倍。

Comments Accepted by SIGMOD 2027

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.10913 2026-06-25 cs.AI cs.PL cs.SE 版本更新

Shepherd: Enabling Programmable Meta-Agents via Reversible Agentic Execution Traces

Shepherd: 一个为元代理提供形式化执行迹的运行时基座

Simon Yu, Derek Chong, Ananjan Nandi, Dilara Soylu, Jiuding Sun, Christopher D Manning, Weiyan Shi

机构 * Northeastern University(东北大学) Stanford University(斯坦福大学)

AI总结 提出Shepherd,一个基于函数式编程的Python运行时基座,将代理执行作为一等对象,通过类似Git的执行迹支持元代理的检查、分叉和重放,在三个用例中显著提升性能。

Comments 50 pages, 22 figures, 14 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.19305 2026-06-25 cs.RO cs.AI cs.CV 版本更新

PhyGile: Physics-Prefix Guided Motion Generation for Agile General Humanoid Motion Tracking

PhyGile: 面向敏捷通用人形运动跟踪的物理前缀引导运动生成

Jiacheng Bao, Haoran Yang, Yucheng Xin, Junhong Liu, Yuecheng Xu, Han Liang, Pengfei Han, Xiaoguang Ma, Dong Wang, Bin Zhao

机构 * Northwestern Polytechnical University(西北工业大学) Shanghai AI Laboratory(上海人工智能实验室) University of Science and Technology of China(中国科学技术大学) Tsinghua University(清华大学) Fudan University(复旦大学) ByteDance(字节跳动) Northeastern University(东北大学)

AI总结 提出PhyGile框架,通过物理前缀引导的机器人原生运动生成,结合课程混合专家训练和后训练,解决文本生成运动在机器人上物理不可行的问题,实现敏捷全身运动跟踪。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.23147 2026-06-25 cs.LG cs.AI 版本更新

Securing Time Integrity in Energy IoT Against Clock Drift and Y2K38 Failures

保护能源物联网中的时间完整性:应对时钟漂移和Y2K38故障

Saeid Jamshidi, Omar Abdul Wahab, Rolando Herrero, Foutse Khomh

机构 * Department of Computer and Software Engineering, Polytechnique Montréal(Polytechnique Montréal 计算机与软件工程系) SWAT Laboratory, Polytechnique Montréal(Polytechnique Montréal SWAT 实验室) College of Engineering, Northeastern University(Northeastern University 工程学院) Concordia Institute for Information Systems Engineering (CIISE), Concordia University(Concordia University 信息系统工程研究所)

AI总结 提出STGAT框架,结合漂移感知时间嵌入、时间自注意力和图注意力,检测能源物联网中的时钟漂移、同步偏移和Y2K38溢出等时间异常,准确率达95.7%。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.23842 2026-06-25 cs.CL 版本更新

How Pragmatics Shape Articulation: A Computational Case Study in STEM ASL Discourse

语用如何塑造发音:STEM ASL 话语的计算案例研究

Saki Imai, Lee Kezar, Laurel Aichler, Mert Inan, Erin Walker, Alicia Wooten, Lorna Quandt, Malihe Alikhani

机构 * Northeastern University(东北大学) Gallaudet University(盖尔道特大学) University of Pittsburgh(匹兹堡大学)

AI总结 通过采集 STEM 领域美国手语对话的运动捕捉数据,量化分析对话、独白和翻译场景中手语发音的时空变化,发现对话中手语持续时间显著缩短,并评估了手语嵌入模型对 STEM 手语的识别能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.24855 2026-06-24 cs.AI 新提交

OpenThoughts-Agent: Data Recipes for Agentic Models

OpenThoughts-Agent: 智能体模型的数据配方

Negin Raoof, Richard Zhuang, Marianna Nezhurina, Etash Guha, Atula Tejaswi, Ryan Marten, Charlie F. Ruan, Tyler Griggs, Alexander Glenn Shaw, Hritik Bansal, E. Kelly Buchanan, Artem Gazizov, Reinhard Heckel, Chinmay Hegde, Sankalp Jajee, Daanish Khazi, Emmanouil Koukoumidis, Xiangyi Li, Hange Liu, Shlok Natarajan, Harsh Raj, Nicholas Roberts, Ethan Shen, Nishad Singhi, Michael Siu, Ashima Suvarna, Hanwen Xing, Patrick Yubeaton, Robert Zhang, Leon Liangyu Chen, Xiaokun Chen, Steven Dillmann, Saadia Gabriel, Xunyi Jiang, Anurag Kashyap, Boxuan Li, Yein Park, Minh Pham, Sujay Sanghavi, Lin Shi, Ke Sun, Yixin Wang, Zhiwei Xu, Erica Zhang, Siyan Zhao, Wanjia Zhao, Jenia Jitsev, Alex Dimakis, Benjamin Feuer, Ludwig Schmidt

机构 * UC Berkeley(加州大学伯克利分校) Stanford University(斯坦福大学) JSC(于利希超级计算中心) LAION University of Texas at Austin(德克萨斯大学奥斯汀分校) Bespoke Labs Laude Institute UCLA(加州大学洛杉矶分校) Harvard University & Harvard Medical School(哈佛大学与哈佛医学院) TU Munich & Munich Center for Machine Learning(慕尼黑工业大学与慕尼黑机器学习中心) New York University(纽约大学) Medical University of South Carolina(南卡罗来纳医科大学) The LLM Data Company BenchFlow Independent Researcher(独立研究员) Northeastern University(东北大学) University of Wisconsin–Madison(威斯康星大学麦迪逊分校) University of Washington(华盛顿大学) TU Darmstadt(达姆施塔特工业大学) University of Southern California(南加州大学) UC San Diego(加州大学圣地亚哥分校) Amazon(亚马逊) Microsoft(微软) Korea University(高丽大学) Cornell Tech(康奈尔科技) University of Michigan(密歇根大学)

AI总结 提出全开放数据筛选流水线,通过100多次消融实验研究任务来源与多样性,构建10万样本训练集,在7个智能体基准上平均44.8%准确率,较最强开源模型提升3.9个百分点。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.24381 2026-06-24 cs.CL cs.AI 新提交

On the Stability of Prompt Ranking in Large Language Model Evaluation

论大语言模型评估中提示排序的稳定性

Shaoshuai Du, Penghao Liang, Yixian Shen, Chuanqi Shi, Hang Zhang, Lun Wang

机构 * University of Amsterdam(阿姆斯特丹大学) Northeastern University(东北大学) University of California San Diego(加州大学圣迭戈分校) Duke University(杜克大学)

AI总结 研究提示排序在常见评估变化下的稳定性,发现顶级提示常变导致选择不可靠,提出基于置信下限的稳定性感知选择策略以提高鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.23939 2026-06-24 math.OC cs.LG cs.NA math.NA 新提交

Constrained Variable Projection for Structured Problems

结构化问题的约束变量投影

Emanuele Zangrando, Sara Venturini, Francesco Rinaldi, Francesco Tudisco

机构 * Gran Sasso Science Institute(格兰萨索科学研究所) MOBS Lab(MOBS实验室) Northeastern University(东北大学) University of Padova(帕多瓦大学) University of Edinburgh(爱丁堡大学)

AI总结 提出约束变量投影框架,将变量投影推广到带凸约束的分离非线性最小二乘问题,推导精确降梯度公式并设计条件梯度算法,在稀疏自编码等任务上提升效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.05174 2026-06-24 cs.MA cs.AI

Emergent Coordination in Multi-Agent Language Models

多智能体语言模型中的涌现协调

Christoph Riedl

机构 * D’Amore-McKim School of Business(达莫-麦克金商学院) Khoury College of Computer Sciences(科里学院计算机科学学院) Network Science Institute(网络科学研究所) Northeastern University(东北大学)

AI总结 研究探讨多智能体系统是否形成更高层次结构,提出信息论框架通过数据驱动方法检测动态涌现,验证身份差异化与目标互补性,揭示交互模式与集体智慧原理的关联。

Journal ref International Conference on Learning Representations (ICLR 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.14862 2026-06-24 cs.CV 版本更新

Fine-Grained Open-Vocabulary Object Detection with Fined-Grained Prompts: Task, Dataset and Benchmark

细粒度开放词汇目标检测与细粒度提示:任务、数据集与基准

Ying Liu, Yijing Hua, Haojiang Chai, Yanbo Wang, TengQi Ye

机构 * department of software engineering, Northeastern University, China(软件工程系,东北大学,中国)

AI总结 本文提出3F-OVD任务,扩展细粒度监督目标检测至开放词汇场景,引入NEU-171K数据集,并提出简单有效的后处理技术。

Comments 8 pages, 4 figures, 2025 IEEE International Conference on Robotics and Automation (ICRA)

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.04477 2026-06-23 cs.LG 版本更新

Data-dependent Exploration for Online Reinforcement Learning from Human Feedback

基于数据的探索:面向人类反馈的在线强化学习

Zhen-Yu Zhang, Yuting Tang, Jiandong Zhang, Lanjihong Ma, Masashi Sugiyama

机构 * Center for Advanced Intelligence Project, RIKEN(日本理化学研究所先进智能研究中心) Graduate School of Frontier Sciences, The University of Tokyo(东京大学前沿科学研究生院) Northeastern University(东北大学) Zhejiang Gongshang University(浙江工商大学)

AI总结 提出数据依赖探索的偏好优化方法(DEPO),利用历史数据构建不确定性奖励,鼓励探索高价值区域,理论证明其数据依赖的遗憾界更紧,实验表明样本效率提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.23590 2026-06-23 cs.AI 新提交

The Topology of Ill-Posed Questions: Persistent Homology for Detection and Steering in LLMs

病态问题的拓扑:用于大语言模型中检测与引导的持续同调

Guangyu Jiang, Sizhe Tang, Mahdi Imani, Tian Lan

机构 * The George Washington University(乔治华盛顿大学) Northeastern University(东北大学)

AI总结 利用持续同调分析LLM内部状态的拓扑结构,统一表示多种病态问题,并实现分类与激活引导,提升响应质量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.22864 2026-06-23 cs.LG 新提交

When AUC 0.998 Is Not Enough: A Candidate Evaluation Protocol for Hidden-State Probes of Indirect Prompt Injection in Multimodal Computer-Use Agents

当AUC 0.998还不够:多模态计算机使用智能体中隐藏状态探针对间接提示注入的候选评估协议

Yanhang Li, Zhichao Fan, Zexin Zhuang

机构 * Northeastern University(东北大学) University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Southern Methodist University(南卫理公会大学)

AI总结 本文通过单骨干案例研究,论证高AUC不能直接证明恶意内容检测,提出后验诊断和候选控制集来明确探针能力的边界。

Comments 17 pages, 3 figures. Camera-ready version for EvalMG '26, The 2nd Workshop on Evaluation for Multimodal Generation, co-located with SIGIR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.22862 2026-06-23 cs.CV cs.LG 新提交

Chains That See, Answers That Don't: A Multi-Aspect Evaluation Recipe for Forced Chain-of-Thought on Video-MME

看得见的链,答不出的答案:针对视频MME上强制思维链的多方面评估方案

Zhichao Fan, Yanhang Li, Zexin Zhuang

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Northeastern University(东北大学) Southern Methodist University(南卫理公会大学)

AI总结 提出三探针评估方案检验强制思维链在视频问答中的可靠性,应用于Qwen2.5-VL发现思维链强依赖视频但未提升准确率,甚至在小模型上导致下降。

Comments 10 pages, 5 figures. To appear at The 2nd Workshop on Evaluation for Multimodal Generation @ SIGIR 2026 (EvalMG '26)

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.22613 2026-06-23 cs.AI 新提交

SkillAudit: From Fixed-Suite Benchmarking to Skill-Centered Assessment

SkillAudit:从固定套件基准测试到以技能为中心的评估

Dexu Yu, Youhua Li, Zhaoyang Guan, Xianhao Lin, Jining Luan, Zihao Rao, Xuanqi Lan, Yang Ran, Bo Lan, Nai-Xin Zhai, Hanwen Du, Junchen Fu, Wenhao Deng, Yongxin Ni, Chunxiao Li

机构 * Northeastern University(东北大学) City University of Hong Kong(香港城市大学) Northwestern University(西北大学) Fudan University(复旦大学) University of Science and Technology of China(中国科学技术大学) Santa Clara University(圣克拉拉大学) Fenz AI Ohio State University(俄亥俄州立大学) University of Glasgow(格拉斯哥大学) National University of Singapore(新加坡国立大学) DeciLix Lab(DeciLix实验室)

AI总结 提出SkillAudit框架,自动生成技能的多维度评估报告,涵盖效用、效率/成本和安全性,通过基线比较和两阶段检测解决固定套件评估的不足。

Comments Preprint. Project page: https://skillaudit.github.io/. Code and evaluation artifacts: https://github.com/SkillAudit/skillaudit

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.22570 2026-06-23 cs.CL 新提交

What are Key Factors for Updates in RL for LLM Reasoning?

RL提升LLM推理能力的关键更新因素是什么?

Peidong Wang, Demi Wang, Xufang Luo, Jiahang Xu, Xiaocui Yang, Shi Feng, Yuqing Yang, Dongsheng Li

机构 * School of Computer Science and Engineering, Northeastern University(东北大学计算机科学与工程学院) Microsoft Research(微软研究院) Carnegie Mellon University(卡内基梅隆大学)

AI总结 通过理论分析RLVR更新,发现离策略程度影响重要性采样比率分布和裁剪行为,提出自适应裁剪策略优化(ACPO),在多种推理基准上优于DAPO和CISPO。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.22307 2026-06-23 cs.LG cs.AI 新提交

Enhancing Protein Representation Learning via Manifold Restore Mixing

通过流形恢复混合增强蛋白质表示学习

Yizhou Dang, Chuang Zhao, Lianbo Ma, Guibing Guo, Xingwei Wang, Zhu Sun

机构 * Software College, Northeastern University(东北大学软件学院) Tianjin University(天津大学) School of Computer Science and Engineering, Northeastern University(东北大学计算机科学与工程学院) Information Systems Technology and Design, Singapore University of Technology and Design(新加坡科技设计大学信息系统技术与设计)

AI总结 针对数据增强破坏蛋白质结构的问题,提出流形恢复混合方法,通过混合原始与增强数据的隐表示恢复结构信息,并引入难度调度器逐步增加训练难度,提升表示学习性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.21345 2026-06-23 cs.CL 新提交

Factual Retrieval in LLMs Is a Redundant, Distributed and Non-Contiguous Process

大型语言模型中的事实检索是一个冗余、分布且非连续的过程

Hail Hochman, Natalie Shapira, Yoav Goldberg

机构 * Bar-Ilan University(巴伊兰大学) Northeastern University(东北大学)

AI总结 本文通过属性计算路径分析,发现LLM中事实检索路径非连续、存在多条功能等价路径,表明知识计算高度冗余和分布。

Comments Accepted to ACL 2026 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.21130 2026-06-23 cs.AI 新提交

Learning Burst-Aware Early Warning Models for Capacity Stress under AI Workload Surges in Hyperscale Data Centers

面向超大规模数据中心AI工作负载激增的突发感知容量压力预警模型学习

Zihan Yu, Xianling Zeng, Zhiming Xue, Yalun Qi, Sichen Zhao

机构 * Northeastern University(东北大学)

AI总结 针对AI工作负载突发导致容量压力的问题,提出部署导向的突发感知预警框架,采用轻量级树模型实现高召回预测,支持主动干预。

详情

展开后加载摘要…

URL PDF HTML 收藏