arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

University of Washington(华盛顿大学)

2026-04-16 至 2026-04-16 共收录 3
2604.14140 2026-04-16 cs.LG cs.AI

LongCoT: Benchmarking Long-Horizon Chain-of-Thought Reasoning

LongCoT:长周期链式推理基准测试

Sumeet Ramesh Motwani, Daniel Nichols, Charles London, Peggy Li, Fabio Pizzati, Acer Blake, Hasan Hammoud, Tavish McDonald, Akshat Naik, Alesia Ivanova, Vignesh Baskaran, Ivan Laptev, Ruben Glatt, Tal Ben-Nun, Philip Torr, Natasha Jaques, Ameya Prabhu, Brian Bartoldson, Bhavya Kailkhura, Christian Schroeder de Witt

机构 * University of Oxford(牛津大学) Lawrence Livermore National Laboratory (LLNL)(劳伦斯利弗莫尔国家实验室) University of Washington(华盛顿大学)

AI总结 LongCoT通过2500个专家设计的问题评估长周期链式推理能力,揭示前沿模型在长时间推理中的不足。

Comments Long-Horizon Reasoning Benchmark

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.14118 2026-04-16 cs.LG math.SP

Complex Interpolation of Matrices with an application to Multi-Manifold Learning

矩阵的复插值及其在多流形学习中的应用

Adi Arbel, Stefan Steinerberger, Ronen Talmon

机构 * Viterbi Faculty of Electrical and Computer Engineering, Technion – Israel Institute of Technology(电气与计算机工程学院,技术学院–以色列理工学院) Department of Mathematics and Department of Applied Mathematics, University of Washington(数学系和应用数学系,华盛顿大学) University of Washington, Seattle, WA 98195, USA(华盛顿大学,西雅图,华盛顿州98195,美国)

AI总结 研究对称正定矩阵插值的谱性质,揭示共同结构与主成分对齐的理论依据,为多流形学习提供方法支持。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.13128 2026-04-16 cs.MA cs.LG cs.RO cs.SY eess.SY

Learning Probabilistic Responsibility Allocations for Multi-Agent Interactions

多智能体交互中的概率责任分配学习

Isaac Remy, Caleb Chang, Karen Leung

机构 * University of Washington, Department of Aeronautics and Astronautics(华盛顿大学航空航天系) NVIDIA

AI总结 本文提出一种学习多智能体交互中概率责任分配模型的方法,通过条件变分自动编码器的潜在空间和多智能体轨迹预测技术,实现基于场景和智能体上下文的责任分配分布学习,展示了在INTERACTION驾驶数据集上的强预测性能和可解释性。

详情

展开后加载摘要…

URL PDF HTML 收藏