arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

Northeastern University(东北大学)

2026-06-01 至 2026-06-01 共收录 7
2605.30621 2026-06-01 cs.AI

Harness Updating Is Not Harness Benefit: Disentangling Evolution Capabilities in Self-Evolving LLM Agents

利用更新并非利用收益:解构自演化LLM智能体中的演化能力

Minhua Lin, Juncheng Wu, Zijun Wang, Zhan Shi, Yisi Sang, Bing He, Zewen Liu, Tianxin Wei, Zongyu Wu, Zhiwei Zhang, Dakuo Wang, Xiang Zhang, Benoit Dumoulin, Cihang Xie, Yuyin Zhou, Suhang Wang, Hanqing Lu

机构 * The Pennsylvania State University(宾夕法尼亚州立大学) UC Santa Cruz(加州大学圣克鲁兹分校) Amazon(亚马逊) Emory University(埃默里大学) UIUC(伊利诺伊大学香槟分校) Northeastern University(东北大学)

AI总结 本文通过分析LLM智能体在外部框架(提示、技能、记忆和工具)上的自演化能力,发现框架更新能力与基础能力无关,而框架收益能力与基础能力呈非单调关系,中等能力模型受益最大。

Comments 24 pages, 9 figures, 12 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.30613 2026-06-01 cs.CR cs.LG

CacheProbe: Auditing Prompt Cache Isolation in Gateway APIs

CacheProbe: 审计网关API中的提示缓存隔离

Ryan Fahey

机构 * Khoury College of Computer Sciences(计算机科学学院) Northeastern University(东北大学)

AI总结 本文通过CacheProbe方法审计OpenRouter API网关架构,发现其共享组织凭证的路由机制可能绕过提供商级别的提示缓存隔离,导致全局缓存共享漏洞。

Comments 11 pages, 8 figures, 2 tables Accepted at SAGAI '26 (Workshop on Secure Agents for Generative AI), co-located with IEEE Symposium on Security and Privacy 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.30519 2026-06-01 cs.CV

OmniMem: Scalable and Adaptive Memory Retrieval for Long Video Generation

OmniMem: 用于长视频生成的可扩展自适应记忆检索

Lin Zhao, Yushu Wu, Yifan Gong, Yanzhi Wang, Pu Zhao

机构 * Northeastern University(东北大学) Adobe Research

AI总结 提出OmniMem框架,通过自适应窗口排除和查询共享KV选择等机制,在自回归视频生成中实现显式全范围稀疏KV检索,显著提升长视频动态程度并保持一致性。

Comments 22 pages, 14 figures; project page: https://wuyushuwys.github.io/OmniMem/

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.29268 2026-06-01 cs.CL cs.AI cs.LG cs.NE

Compute Allocation in Evolutionary Search: From Depth-Breadth to Multi-Armed Bandits

进化搜索中的计算分配:从深度-广度到多臂老虎机

Sixue Xing, Haoyu He, Kerui Wu, Zhuo Yang, Haozheng Luo, Tianfan Fu, Aarthy Nagarajan

机构 * University of Notre Dame(诺丁汉大学) Northeastern University(东北大学) University of Massachusetts Amherst(马萨诸塞大学阿默斯特分校) Southeast University(东南大学) Northwestern University(西北大学) Nanjing University(南京大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)

AI总结 针对LLM引导的进化搜索中固定预算的LLM调用分配问题,提出基于多臂老虎机的BaSE方法,通过跨并行轨迹分配调用,平均适应度提升12.3%。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.16084 2026-06-01 cs.CV cs.AI

SpectralTrain: A Universal Framework for Hyperspectral Image Classification

SpectralTrain:一种通用的高光谱图像分类框架

Meihua Zhou, Liping Yu, Xinyu Tong, Wai Kin Fung, Ruiguo Hu, Jiarui Zhao, Nan Wan

机构 * School of Medical Information, Wannan Medical University(皖南医学院信息学院) University of Chinese Academy of Sciences(中国科学院大学) The Chinese University of Hong Kong(香港中文大学) Northeastern University(东北大学)

AI总结 提出SpectralTrain通用训练框架,通过课程学习与基于PCA的光谱下采样提升高光谱图像分类效率,在多个数据集上实现2-7倍训练加速且精度损失小。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.08964 2026-06-01 cs.LG cs.AI cs.CL cs.CY

A Behavioural and Representational Evaluation of Goal-Directedness in Language Model Agents

语言模型智能体中目标导向性的行为与表征评估

Raghu Arghal, Fade Chen, Niall Dalton, Evgenii Kortukov, Calum McNamara, Angelos Nalmpantis, Moksh Nirvaan, Gabriele Sarti, Mario Giulianelli

机构 * University of Pennsylvania(宾夕法尼亚大学) New York University(纽约大学) Indiana University, Bloomington(印第安纳大学,布卢明顿) Northeastern University(东北大学) University College London(伦敦大学学院)

AI总结 本文提出一种结合行为评估与内部表征可解释性分析的目标导向性评估框架,并以LLM智能体在2D网格世界中的导航为例,验证了其行为与表征的一致性。

Comments Proceedings of the 43rd International Conference on Machine Learning (ICML 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.14583 2026-06-01 cs.AI

LLM Bias Evaluation: Gender, Racial, and Age Disparities in Occupational and Crime Scenarios

LLM偏差评估:职业与犯罪场景中的性别、种族和年龄差异

Vishal Mirza, Rahul Kulkarni, Aakanksha Jadhav

机构 * New York University(纽约大学) Northeastern University(东北大学) Washington University in St. Louis(圣路易斯华盛顿大学)

AI总结 本文评估了2024年四大领先LLM在职业和犯罪场景中的性别、种族和年龄偏差,发现去偏努力常导致新的公平性权衡,即“去偏悖论”。

Comments Updated title and abstract to emphasize key findings on the debiasing paradox for improved discoverability. Content and findings unchanged. 11 pages, 17 figures, Accepted at IEEE Conference on Artificial Intelligence (IEEE CAI) 2025. Full Paper acceptance in the Vertical HUMAN-CENTERED AI category

Journal ref 2025 IEEE Conference on Artificial Intelligence (CAI)

详情

展开后加载摘要…

URL PDF HTML 收藏