arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

Northeastern University(东北大学)

2026-06-23 至 2026-06-23 共收录 10
2606.23590 2026-06-23 cs.AI 新提交

The Topology of Ill-Posed Questions: Persistent Homology for Detection and Steering in LLMs

病态问题的拓扑:用于大语言模型中检测与引导的持续同调

Guangyu Jiang, Sizhe Tang, Mahdi Imani, Tian Lan

机构 * The George Washington University(乔治华盛顿大学) Northeastern University(东北大学)

AI总结 利用持续同调分析LLM内部状态的拓扑结构,统一表示多种病态问题,并实现分类与激活引导,提升响应质量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.22864 2026-06-23 cs.LG 新提交

When AUC 0.998 Is Not Enough: A Candidate Evaluation Protocol for Hidden-State Probes of Indirect Prompt Injection in Multimodal Computer-Use Agents

当AUC 0.998还不够:多模态计算机使用智能体中隐藏状态探针对间接提示注入的候选评估协议

Yanhang Li, Zhichao Fan, Zexin Zhuang

机构 * Northeastern University(东北大学) University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Southern Methodist University(南卫理公会大学)

AI总结 本文通过单骨干案例研究,论证高AUC不能直接证明恶意内容检测,提出后验诊断和候选控制集来明确探针能力的边界。

Comments 17 pages, 3 figures. Camera-ready version for EvalMG '26, The 2nd Workshop on Evaluation for Multimodal Generation, co-located with SIGIR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.22862 2026-06-23 cs.CV cs.LG 新提交

Chains That See, Answers That Don't: A Multi-Aspect Evaluation Recipe for Forced Chain-of-Thought on Video-MME

看得见的链,答不出的答案:针对视频MME上强制思维链的多方面评估方案

Zhichao Fan, Yanhang Li, Zexin Zhuang

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Northeastern University(东北大学) Southern Methodist University(南卫理公会大学)

AI总结 提出三探针评估方案检验强制思维链在视频问答中的可靠性,应用于Qwen2.5-VL发现思维链强依赖视频但未提升准确率,甚至在小模型上导致下降。

Comments 10 pages, 5 figures. To appear at The 2nd Workshop on Evaluation for Multimodal Generation @ SIGIR 2026 (EvalMG '26)

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.22613 2026-06-23 cs.AI 新提交

SkillAudit: From Fixed-Suite Benchmarking to Skill-Centered Assessment

SkillAudit:从固定套件基准测试到以技能为中心的评估

Dexu Yu, Youhua Li, Zhaoyang Guan, Xianhao Lin, Jining Luan, Zihao Rao, Xuanqi Lan, Yang Ran, Bo Lan, Nai-Xin Zhai, Hanwen Du, Junchen Fu, Wenhao Deng, Yongxin Ni, Chunxiao Li

机构 * Northeastern University(东北大学) City University of Hong Kong(香港城市大学) Northwestern University(西北大学) Fudan University(复旦大学) University of Science and Technology of China(中国科学技术大学) Santa Clara University(圣克拉拉大学) Fenz AI Ohio State University(俄亥俄州立大学) University of Glasgow(格拉斯哥大学) National University of Singapore(新加坡国立大学) DeciLix Lab(DeciLix实验室)

AI总结 提出SkillAudit框架,自动生成技能的多维度评估报告,涵盖效用、效率/成本和安全性,通过基线比较和两阶段检测解决固定套件评估的不足。

Comments Preprint. Project page: https://skillaudit.github.io/. Code and evaluation artifacts: https://github.com/SkillAudit/skillaudit

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.22570 2026-06-23 cs.CL 新提交

What are Key Factors for Updates in RL for LLM Reasoning?

RL提升LLM推理能力的关键更新因素是什么?

Peidong Wang, Demi Wang, Xufang Luo, Jiahang Xu, Xiaocui Yang, Shi Feng, Yuqing Yang, Dongsheng Li

机构 * School of Computer Science and Engineering, Northeastern University(东北大学计算机科学与工程学院) Microsoft Research(微软研究院) Carnegie Mellon University(卡内基梅隆大学)

AI总结 通过理论分析RLVR更新,发现离策略程度影响重要性采样比率分布和裁剪行为,提出自适应裁剪策略优化(ACPO),在多种推理基准上优于DAPO和CISPO。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.22307 2026-06-23 cs.LG cs.AI 新提交

Enhancing Protein Representation Learning via Manifold Restore Mixing

通过流形恢复混合增强蛋白质表示学习

Yizhou Dang, Chuang Zhao, Lianbo Ma, Guibing Guo, Xingwei Wang, Zhu Sun

机构 * Software College, Northeastern University(东北大学软件学院) Tianjin University(天津大学) School of Computer Science and Engineering, Northeastern University(东北大学计算机科学与工程学院) Information Systems Technology and Design, Singapore University of Technology and Design(新加坡科技设计大学信息系统技术与设计)

AI总结 针对数据增强破坏蛋白质结构的问题,提出流形恢复混合方法,通过混合原始与增强数据的隐表示恢复结构信息,并引入难度调度器逐步增加训练难度,提升表示学习性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.21345 2026-06-23 cs.CL 新提交

Factual Retrieval in LLMs Is a Redundant, Distributed and Non-Contiguous Process

大型语言模型中的事实检索是一个冗余、分布且非连续的过程

Hail Hochman, Natalie Shapira, Yoav Goldberg

机构 * Bar-Ilan University(巴伊兰大学) Northeastern University(东北大学)

AI总结 本文通过属性计算路径分析,发现LLM中事实检索路径非连续、存在多条功能等价路径,表明知识计算高度冗余和分布。

Comments Accepted to ACL 2026 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.21130 2026-06-23 cs.AI 新提交

Learning Burst-Aware Early Warning Models for Capacity Stress under AI Workload Surges in Hyperscale Data Centers

面向超大规模数据中心AI工作负载激增的突发感知容量压力预警模型学习

Zihan Yu, Xianling Zeng, Zhiming Xue, Yalun Qi, Sichen Zhao

机构 * Northeastern University(东北大学)

AI总结 针对AI工作负载突发导致容量压力的问题,提出部署导向的突发感知预警框架,采用轻量级树模型实现高召回预测,支持主动干预。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.20877 2026-06-23 cs.MA cs.AI cs.SI nlin.AO physics.soc-ph 新提交

Artificial collectives of specialists and generalists excel at different tasks

专家与通才的人工集体在不同任务中表现出色

John Meluso, Laurent Hébert-Dufresne, Christoph Riedl, H. Oliver Gao

机构 * Cornell University(康奈尔大学) University of Vermont(佛罗里达大学) Santa Fe Institute(圣达菲研究所) Northeastern University(东北大学)

AI总结 通过优化代理的系统实验,研究代理解释能力、理性边界和任务质量如何影响集体绩效,发现通才集体在生成、选择和协调任务上更优,而专家集体在谈判任务上更优。

Comments 10 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.18123 2026-06-23 cs.CV 新提交

Predicting Immune Biomarkers with MultiModal Mixture-of-Expert Pathology Foundation Models Empowers Precision Oncology

使用多模态混合专家病理基础模型预测免疫生物标志物,赋能精准肿瘤学

Tianyu Liu, Ziqing Wang, Zhaokang Liang, Tong Ding, Peter Humphrey, Lorraine Colón-Cartagena, Emily Ling-Lin Pai, Kenneth Tou En Chang, Mohamed Kahila, Jonathan Chong Kai Liew, Tinglin Huang, Rex Ying, Kaize Ding, Faisal Mahmood, Wengong Jin

机构 * Program of Computational Biology and Bioinforamtics, Yale University(耶鲁大学计算生物学与生物信息学项目) Broad Institute of MIT and Harvard(麻省理工学院与哈佛大学博德研究所) Department of Statistics and Data Science, Northwestern University(西北大学统计与数据科学系) Department of Computer Science, Northeastern University(东北大学计算机科学系) Department of Computer Science, Harvard University(哈佛大学计算机科学系) Department of Pathology, Yale University(耶鲁大学病理学系) Department of Anatomic Pathology and Laboratory Medicine, Hospital of the University of Pennsylvania(宾夕法尼亚大学医院解剖病理学与检验医学系) Department of Pathology and Laboratory Medicine, University of California, San Francisco(加州大学旧金山分校病理学与检验医学系) Department of Pathology and Laboratory Medicine, KK Women’s and Children’s Hospital(竹脚妇幼医院病理学与检验医学系) Department of Biostatistics, Epidemiology and Informatics, Perelman School of Medicine, University of Pennsylvania(宾夕法尼亚大学佩雷尔曼医学院生物统计学、流行病学与信息学系)

AI总结 提出MixTIME多模态基础模型,采用混合专家架构整合不同模态的病理基础模型,从HE全切片图像预测多重免疫荧光蛋白表达,在17个蛋白标记物上达到最优性能,并增强空间域识别、生存预测等下游任务。

Comments 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏