arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

大模型推理能力

大模型数学、逻辑、规划、多步推理和测试时计算能力。

2026-03-09 至 2026-03-09 共收录 14 信号源:cs.CL, cs.AI, cs.LG

1. 规划推理 14 篇

2603.05818 2026-03-09 cs.CL 87%

RouteGoT: Node-Adaptive Routing for Cost-Efficient Graph of Thoughts Reasoning

RouteGoT: 用于图状思维推理的节点自适应路由

Yuhang Liu, Ruijie Wang, Yunlong Chu, Bing Hao, Yumeng Lin, Shengzhong Liu, Minglai Shao

机构 * School of New Media and Communication, Tianjin University(新媒体与传播学院,天津大学) School of Computer Science and Engineering, Beihang University(计算机科学与工程学院,北航) Shanghai Jiao Tong University(上海交通大学)

专题命中 规划推理 :reasoning(title,abstract);chain-of-thought(abstract);CoT(abstract);planning(abstract)

AI总结 RouteGoT通过节点自适应路由框架提升图状思维推理的效率与准确性,实现性能与成本的平衡。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.24142 2026-03-09 cs.CL cs.AI 86%

CoME: Empowering Channel-of-Mobile-Experts with Informative Hybrid-Capabilities Reasoning

CoME:赋能移动专家的 informative hybrid-capabilities 推理

Yuxuan Liu, Weikai Xu, Kun Huang, Changyu Chen, Jiankun Zhao, Pengzhi Gao, Wei Liu, Jian Luan, Shuo Shang, Bo Du, Ji-Rong Wen, Rui Yan

机构 * ruc(中国人民大学人工智能学院) xiaomi(小米公司) ntu(南洋理工大学) whu(武汉大学)

专题命中 规划推理 :reasoning(title,abstract);CoT(abstract);planning(abstract);分类 cs.CL、cs.AI

AI总结 CoME 通过四种专家架构和渐进训练策略,实现移动代理混合能力推理的解耦增强与平衡优化。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.03704 2026-03-09 cs.RO cs.AI 83%

Large-Language-Model-Guided State Estimation for Partially Observable Task and Motion Planning

基于大语言模型的态估计用于部分可观测任务和运动规划

Yoonwoo Kim, Raghav Arora, Roberto Martín-Martín, Peter Stone, Ben Abbatematteo, Yoonchang Sung

机构 * The University of Texas at Austin(德克萨斯大学) Nanyang Technological University(南洋理工大学)

专题命中 规划推理 :planning(title,abstract);reasoning(abstract);分类 cs.AI

AI总结 本文提出CoCo-TAMP框架,利用大语言模型的常识推理能力,通过分层态估计提升部分可观测任务和运动规划的效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06338 2026-03-09 eess.IV cs.AI cs.LG cs.SY eess.SY physics.med-ph 81%

AI End-to-End Radiation Treatment Planning Under One Second

AI端到端放射治疗计划在1秒内完成

Simon Arberet, Riqiang Gao, Martin Kraus, Florin C. Ghesu, Wilko Verbakel, Mamadou Diallo, Anthony Magliari, Venkatesan Karuppusamy, Sushil Beriwal, REQUITE Consortium, Ali Kamen, Dorin Comaniciu

机构 * Digital Technology and Innovation(数字技术与创新) Siemens Healthineers(西门子医疗) Varian Medical Affairs(Varian医疗事务)

专题命中 规划推理 :planning(title,abstract);分类 cs.AI、cs.LG

AI总结 本文提出了一种端到端深度学习框架AIRT,可在1秒内生成高质量的放射治疗计划,实现了超快速且标准化的治疗计划制定。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06064 2026-03-09 cs.AI 79%

Agentic LLM Planning via Step-Wise PDDL Simulation: An Empirical Characterisation

通过分步PDDL模拟实现代理LLM规划:一项实证分析

Kai Göbel, Pierrick Lorang, Patrik Zips, Tobias Glück

机构 * AIT Austrian Institute of Technology GmbH(奥地利技术研究院)

专题命中 规划推理 :planning(title,abstract);分类 cs.AI

AI总结 本文通过PyPDDLEngine实验证明代理LLM在任务规划中相比传统方法具有小幅优势,但效果受环境反馈类型影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00814 2026-03-09 cs.RO cs.LG cs.SY eess.SY 79%

Real-Time Learning of Predictive Dynamic Obstacle Models for Robotic Motion Planning

实时学习预测性动态障碍物模型用于机器人运动规划

Stella Kombo, Masih Haseli, Skylar X. Wei, Joel W. Burdick

机构 * Division of Engineering and Applied Science, California Institute of Technology(工程与应用科学分校,加州理工学院) Applied Intuition(应用直觉)

专题命中 规划推理 :planning(title,abstract);分类 cs.LG

AI总结 本文提出一种实时学习动态障碍物预测模型的方法,通过改进的Hankel-DMD实现去噪和多步预测,适用于机器人运动规划。

Comments 10 pages, 6 figures, submitted to IEEE International Conference on Robotics and Automation (ICRA) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06190 2026-03-09 cs.RO 78%

DreamToNav: Generalizable Navigation for Robots via Generative Video Planning

DreamToNav: 通过生成式视频规划实现通用机器人导航

Valerii Serpiva, Jeffrin Sam, Chidera Simon, Hajira Amjad, Iana Zhura, Artem Lykov, Dzmitry Tsetserukou

机构 * Intelligent Space Robotics Laboratory, Skolkovo Institute of Science and Technology(智能空间机器人实验室,斯克尔科夫科学与技术研究所)

专题命中 规划推理 :planning(title,abstract)

AI总结 DreamToNav通过生成式视频规划实现通用机器人导航,利用自然语言提示生成精确运动路径,实现目标导向的自主导航。

Comments Submitted to conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06084 2026-03-09 cs.RO 78%

Multimodal Behavior Tree Generation: A Small Vision-Language Model for Robot Task Planning

多模态行为树生成:一种小型视觉-语言模型用于机器人任务规划

Cristiano Battistini, Riccardo Andrea Izzo, Gianluca Bardaro, Matteo Matteucci

机构 * Department of Electronics, Information, and Bioengineering, Politecnico di Milano(电子信息与生物工程系,米兰理工学院)

专题命中 规划推理 :planning(title,abstract)

AI总结 本文提出一种小型视觉-语言模型,通过生成行为树提升机器人任务规划效率,实验证明其在性能和资源消耗上均优于现有闭源模型。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05831 2026-03-09 cs.DC 78%

Knowledge-driven Reasoning for Mobile Agentic AI: Concepts, Approaches, and Directions

基于知识的移动代理AI推理:概念、方法与方向

Guangyuan Liu, Changyuan Zhao, Yinqiu Liu, Dusit Niyato, Biplab Sikdar

专题命中 规划推理 :reasoning(title,abstract)

AI总结 基于知识的移动代理AI推理通过提取可重用决策结构,同步并注入设备端推理以降低延迟、能耗和误差积累,提出DIKW分类法区分不同知识表示形式,并验证了知识暴露的非单调性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05697 2026-03-09 cs.CV 78%

MultiHaystack: Benchmarking Multimodal Retrieval and Reasoning over 40K Images, Videos, and Documents

MultiHaystack:用于40,000张图像、视频和文档的多模态检索和推理的基准测试

Dannong Xu, Zhongyu Yang, Jun Chen, Yingfang Yuan, Ming Hu, Lei Sun, Luc Van Gool, Danda Pani Paudel, Chun-Mei Feng

机构 * INSAIT Lanzhou University(兰州大学) King Abdullah University of Science and Technology(国王阿卜杜勒-阿齐兹科学与技术大学) Heriot-Watt University(赫瑞-沃德大学) Monash University(莫纳什大学) University College Dublin(都柏林大学学院)

专题命中 规划推理 :reasoning(title,abstract)

AI总结 MultiHaystack是首个评估大规模异构多模态检索与推理的基准测试,揭示了MLLMs在异构语料库中检索证据时的性能瓶颈。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05270 2026-03-09 cs.RO cs.AI cs.HC cs.MA cs.SY eess.SY 74%

XR-DT: Extended Reality-Enhanced Digital Twin for Safe Motion Planning via Human-Aware Model Predictive Path Integral Control

XR-DT:增强现实增强型数字孪生用于通过人感知模型预测路径积分控制的安全运动规划

Tianyi Wang, Jiseop Byeon, Ahmad Yehia, Yiming Xu, Jihyung Park, Tianyi Zeng, Sikai Chen, Ziran Wang, Junfeng Jiao, Christian Claudel

机构 * Department of Civil, Architectural, and Environmental Engineering, The University of Texas at Austin(德克萨斯大学奥斯汀分校土木、建筑与环境工程系) School of Architecture, The University of Texas at Austin(德克萨斯大学奥斯汀分校建筑学院) School of Civil and Construction Engineering, Purdue University(普渡大学土木与建设工程学院) Department of Civil and Environmental Engineering, University of Wisconsin-Madison(威斯康星大学麦迪逊分校土木与环境工程系)

专题命中 规划推理 :planning(title);分类 cs.AI

AI总结 XR-DT通过结合增强现实与数字孪生技术,提出HA-MPPI控制模型,实现基于人类行为预测的安全高效人机交互。

Comments 8 pages, 6 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17938 2026-03-09 cs.CL cs.LG 73%

SPINE: Token-Selective Test-Time Reinforcement Learning with Entropy-Band Regularization

SPINE:基于熵带正则化的令牌选择性测试时间强化学习

Jianghao Wu, Yasmeen George, Jin Ye, Yicheng Wu, Daniel F. Schmidt, Jianfei Cai

机构 * Monash University(墨尔本大学) Imperial College London(伦敦帝国理工学院)

专题命中 规划推理 :reasoning(abstract);chain-of-thought(abstract);分类 cs.CL、cs.LG

AI总结 SPINE通过令牌选择性和熵带正则化提升测试时间推理稳定性与效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.14063 2026-03-09 cs.RO 67%

Language Conditioning Improves Accuracy of Aircraft Goal Prediction in Non-Towered Airspace

语言引导提升非塔台空域飞机目标预测的准确性

Sundhar Vinodh Sangeetha, Chih-Yuan Chiu, Sarah H. Q. Li, Shreyas Kousik

机构 * Georgia Institute of Technology(佐治亚理工学院) School of Aerospace Engineering(航空航天工程学院) School of Electrical and Computer Engineering(电气与计算机工程学院) School of Mechanical Engineering(机械工程学院)

专题命中 规划推理 :reasoning(abstract);planning(abstract)

AI总结 本文提出利用语言信息提升非塔台空域飞机目标预测准确性的多模态框架,通过整合自然语言理解和空间推理,有效提高自主决策能力。

Comments The last two authors advised equally. Accepted to the 2026 IEEE International Conference on Robotics and Automation. 8 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23405 2026-03-09 cs.LG 57%

Planner Aware Path Learning in Diffusion Language Models Training

扩散语言模型训练中的规划感知路径学习

Fred Zhangzhi Peng, Zachary Bezemek, Jarrid Rector-Brooks, Shuibai Zhang, Anru R. Zhang, Michael Bronstein, Alexander Tong, Avishek Joey Bose

机构 * Duke University(杜克大学) Mila Université de Montréal(蒙特利尔大学) California Institute of Technology(加州理工学院) University of Wisconsin–Madison(威斯康星大学麦迪逊分校) University of Oxford(牛津大学) AITHYRA Imperial College London(伦敦帝国学院)

专题命中 规划推理 :planning(abstract);分类 cs.LG

AI总结 本文提出PAPL,一种通过规划感知路径学习提升扩散语言模型训练与推理一致性的方法,实现跨领域性能提升。

Comments Camera ready version for ICLR2026

详情

展开后加载摘要…

URL PDF HTML 收藏