arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

University of Michigan(密歇根大学安娜堡分校)

2026-05-11 至 2026-05-11 共收录 14
2605.08060 2026-05-11 cs.CL cs.AI cs.GT cs.MA

The Memory Curse: How Expanded Recall Erodes Cooperative Intent in LLM Agents

记忆诅咒:扩展回忆如何在LLM代理中侵蚀合作意图

Jiayuan Liu, Tianqin Li, Shiyi Du, Xin Luo, Haoxuan Zeng, Emanuel Tewolde, Tai Sing Lee, Tonghan Wang, Carl Kingsford, Vincent Conitzer

机构 * Carnegie Mellon University(卡内基梅隆大学) Foundations of Cooperative AI Lab (FOCAL)(合作人工智能基础实验室) University of Michigan(密歇根大学) Harvard University(哈佛大学)

AI总结 研究发现扩展回忆会系统性地削弱多代理社会困境中的合作意图,通过三种分析揭示记忆内容对合作的影响,证明记忆是影响多代理行为的主动因素。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.07969 2026-05-11 cs.LG cs.IT math.IT

When Diffusion Model Can Ignore Dimension: An Entropy-Based Theory

扩散模型可以忽略维度:基于熵的理论

Ahmad Aghapour, Erhan Bayraktar

机构 * Department of Mathematics, University of Michigan(数学系,密歇根大学)

AI总结 本文从信息论角度探讨扩散采样收敛性,证明高维数据下采样效率由潜在表示熵决定,而非环境维度,为高维空间高效采样提供理论支持。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.07625 2026-05-11 cs.RO cs.SY eess.SY

GATO: GPU-Accelerated and Batched Trajectory Optimization for Scalable Edge Model Predictive Control

GATO:用于可扩展边缘模型预测控制的GPU加速和批处理轨迹优化

Alexander Du, Emre Adabag, Gabriel Bravo-Palacios, Brian Plancher

机构 * School of Engineering and Applied Science, Columbia University(哥伦比亚大学工程与应用科学学院) University of Michigan(密歇根大学) Barnard College, Columbia University and Dartmouth College(哥伦比亚大学巴纳德学院和达特茅斯学院)

AI总结 GATO通过算法、软件和计算硬件的协同设计,实现对中等批量大小的实时轨迹优化,提升边缘模型预测控制的性能。

Comments Accepted to ICRA 2026. 8 pages, 8 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.07331 2026-05-11 cs.LG cs.AI

Rethinking Importance Sampling in LLM Policy Optimization: A Cumulative Token Perspective

重新思考LLM策略优化中的重要性采样:从累积token视角

Yuheng Zhang, Chenlu Ye, Shuowei Jin, Changlong Yu, Wei Xiong, Saurabh Sahu, Nan Jiang

机构 * UIUC(伊利诺伊大学香槟分校) University of Michigan(密歇根大学) Amazon(亚马逊)

AI总结 本文提出CTPO,通过累积token重要性采样比解决偏倚-方差困境,结合位置自适应裁剪提升稳定性,实现更一致的正则化,在数学推理基准上表现最佳。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.07305 2026-05-11 cs.CL cs.AI

MedAction: Towards Active Multi-turn Clinical Diagnostic LLMs

MedAction:迈向主动多轮临床诊断LLMs

Hsin-Ling Hsu, Zizheng Wang, Donghua Zhang, Nai-Chia Chen, Jerry Wang, Jun-En Ding, Chia-Hsuan Hsu, Guoan Wang, Feng Liu, Fang-Ming Hung, Chenwei Wu, Liyue Shen

机构 * National Chengchi University(国立中正大学) Georgetown University(乔治城大学) University of Michigan(密歇根大学) Stevens Institute of Technology(史蒂文斯理工学院) National Taiwan University of Science and Technology(台湾科技大学) Far Eastern Memorial Hospital(东方纪念医院)

AI总结 本文提出MedAction,通过LLM与环境交互生成高质量多轮诊断轨迹,解决现有模型在动态证据下推理不足的问题,提升开源医疗LLMs性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.07180 2026-05-11 cs.CL

Learning Agent Routing From Early Experience

从早期经验学习代理路由

Yimin Wang, Jiahao Qiu, Xuan Qi, Xinzhe Juan, Jingzhe Shi, Zelin Zhao, Hongru Wang, Shilong Liu, Mengdi Wang

机构 * AI Lab, Princeton University(普林斯顿大学人工智能实验室) University of Michigan(密歇根大学) Institute for Interdisciplinary Information Sciences (IIIS), Tsinghua University(清华大学交叉信息学院) Shanghai Jiao Tong University(上海交通大学) University of Edinburgh(爱丁堡大学) King’s College London(伦敦国王学院)

AI总结 本文研究在冷启动条件下如何在轻量级LLM推理与完整代理执行间进行路由,提出无需训练的BoundaryRouter框架,通过早期行为经验和指导性推理决定是否直接使用LLM还是升级至代理,实验显示其在降低推理时间与提升性能方面优于其他方法。

Comments 17 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.07139 2026-05-11 cs.CL cs.AI cs.LG

Structural Rationale Distillation via Reasoning Space Compression

通过推理空间压缩实现结构性理性提炼

Jialin Yang, Jiankun Wang, Jiajun Wu, Henry Leung, Jiayu Zhou, Steve Drew

机构 * University of Calgary(卡尔加里大学) University of Michigan(密歇根大学)

AI总结 本文提出D-RPC方法,通过压缩推理空间约束教师模型生成一致且多样化的推理路径,提升学生模型在数学和常识推理任务中的表现,优于多种基线方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.07063 2026-05-11 cs.LG cs.AI

Dr. Post-Training: A Data Regularization Perspective on LLM Post-Training

Dr. Post-Training:一种数据正则化视角下的LLM后训练

Pingbang Hu, Xueshen Liu, Z. Morley Mao, Jiaqi W. Ma

机构 * University of Illinois Urbana–Champaign(伊利诺伊大学厄巴纳-香槟分校) University of Michigan(密歇根大学)

AI总结 本文提出Dr. Post-Training框架,通过将通用训练数据作为数据诱导正则化器,防止模型过拟合稀缺目标数据,从而提升LLM后训练效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.07042 2026-05-11 cs.AI cs.LG

The Context Gathering Decision Process: A POMDP Framework for Agentic Search

上下文收集决策过程:一种用于代理搜索的POMDP框架

Chinmaya Kausik, Adith Swaminathan, Nathan Kallus

机构 * University of Michigan(密歇根大学) Netflix

AI总结 本文提出CGDP框架,通过将LLM行为建模为近似汤普森采样,引入谓词方法分解搜索过程,并设计两种干预措施提升多跳推理能力。

Comments 25 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.06913 2026-05-11 astro-ph.EP astro-ph.IM cs.LG

You Only Stack Once (YOSO): A Motion-Filtered, Deep-Learning Framework for Detecting Faint Moving Sources

你只堆叠一次(YOSO):一种运动过滤的深度学习框架,用于检测微弱移动源

Nitya Pandey, César Fuentes, Pedro Bernardinelli, Valeria Frías, Colin Orion Chandler, David E. Trilling, Matthew J. Holman, Steven Stetzler, Dallin Spencer, Hsing Wen Lin, Luis E. Salazar Manzano, Darin Ragozzine, Ryder Strauss, Mario Jurić, Andrew J. Connolly, Hayden Smotherman, Scott S. Sheppard, Kevin Napier

机构 * Dept. of Astronomy \& the DiRAC Institute, University of Washington, Seattle, USA Facultad de Ciencias Físicas y Matemáticas (FCFM), University of Chile, Beauchef 850, 851, Santiago, Chile LSST Interdisciplinary Network for Collaboration Department of Astronomy Planetary Science, Northern Arizona University, Flagstaff, USA Harvard-Smithsonian Center for Astrophysics, 60 Garden Street, MS 51, Cambridge, MA 02138, USA Jet Propulsion Laboratory, California Institute of Technology, 4800 Oak Grove Dr., Pasadena, CA 91109 USA Brigham Young University, Department of Physics Department of Physics, University of Michigan, Ann Arbor, MI 48109, USA Michigan Institute for Data AI in Society, University of Michigan, Ann Arbor, MI 48109, USA Department of Astronomy, University of Michigan, Ann Arbor, MI 48109, USA eScience Institute, Department of Astronomy, University of Washington, Seattle, WA 98195-1580, USA Planets Laboratory, Carnegie Institution for Science, Washington, DC 20015

AI总结 YOSO通过运动过滤技术检测宽视场天文调查中的微弱慢速太阳系天体,其核心方法是Gaussian Motion Filter,能有效提升信噪比,发现45个已知天体和11个新冥王星特异天体,适用于大规模调查及行星成像等领域。

Comments Accepted to The Astronomical Journal; 13 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.06863 2026-05-11 cs.RO cs.HC

Bi3: A Biplatform, Bicultural, Biperson Dataset for Social Robot Navigation

Bi3:一个双平台、双文化、双人数据集用于社交机器人导航

Andrew Stratton, Phani Teja Singamaneni, Pranav Goyal, Rachid Alami, Christoforos Mavrogiannis

机构 * Department of Robotics, University of Michigan(机器人系,密歇根大学) LAAS-CNRS, University of Toulouse(LAAS-CNRS,图卢兹大学) INRIA, University of Lorraine(INRIA,洛林大学)

AI总结 Bi3数据集通过双平台、双文化、双人交互,为社交机器人导航提供多样性模型复杂性的基准,用于研究人类与机器人在受限环境中的协同活动。

Comments ICRA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.06756 2026-05-11 cs.LG cs.SY eess.SY

Physics-based Digital Twins for Integrated Thermal Energy Systems Using Active Learning

基于物理的数字孪生用于集成热能系统的主动学习

Umme Mahbuba Nabila, Paul Seurin, Linyu Lin, Majdi I. Radaideh

机构 * a Department of Nuclear Engineering Radiological Sciences, University of Michigan, Ann Arbor, MI 48109, United States b Department of Computer Science Engineering, University of Michigan, Ann Arbor, MI 48109, United States c Nuclear Science \& Technology Division, Idaho National Laboratory, Idaho Falls, ID 83415, United States

AI总结 本文提出结合物理模型与数据驱动方法的主动学习框架,用于高效准确地控制热能分布系统,通过减少模拟轨迹数量提升预测精度和计算效率。

Comments 23 pages, 12 figures, and 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.02881 2026-05-11 cs.RO

MolmoAct2: Action Reasoning Models for Real-world Deployment

MolmoAct2:面向现实部署的动作推理模型

Haoquan Fang, Jiafei Duan, Donovan Clay, Sam Wang, Shuo Liu, Weikai Huang, Xiang Fan, Wei-Chuan Tsai, Shirui Chen, Yi Ru Wang, Shanli Xing, Jaemin Cho, Jae Sung Park, Ainaz Eftekhar, Peter Sushko, Karen Farley, Angad Wadhwa, Cole Harrison, Winson Han, Ying-Chun Lee, Eli VanderBilt, Rose Hendrix, Suveen Ellawela, Lucas Ngoo, Joyce Chai, Zhongzheng Ren, Ali Farhadi, Dieter Fox, Ranjay Krishna

机构 * Allen Institute for AI(艾伦人工智能研究所) University of Washington(华盛顿大学) National University of Singapore(新加坡国立大学) University of Pennsylvania(宾夕法尼亚大学) Johns Hopkins University(约翰霍普金斯大学) Amazon(亚马逊公司) University of Michigan(密歇根大学) University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校)

AI总结 本文提出MolmoAct2,一个完全开放的动作推理模型,通过改进架构和引入新数据集,在五个方面提升性能,展示了在多个基准测试中优于现有模型的成果。

Comments 31 pages, project page: https://allenai.org/blog/molmoact2

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.04183 2026-05-11 cs.CL cs.AI cs.CY cs.HC

Seeing Like an AI: How LLMs Apply (and Misapply) Wikipedia Neutrality Norms

像AI一样看待:LLMs如何应用(并误用)维基百科中立规范

Joshua Ashkinaze, Ruijia Guan, Laura Kurek, Eytan Adar, Ceren Budak, Eric Gilbert

机构 * University of Michigan(密歇根大学)

AI总结 研究评估LLMs在检测和纠正偏见维基百科编辑时的表现,发现其在中立性检测上准确率低,但生成任务表现较好,但存在额外非中立性修改,引发对社区规范执行与公众认知之间差异的思考。

Comments Appeared at ICWSM 2026

详情

展开后加载摘要…

URL PDF HTML 收藏