Dream-MPC: Gradient-Based Model Predictive Control with Latent Imagination
Dream-MPC:基于梯度与潜在想象的模型预测控制
机构 * Autonomous Intelligent Systems, Computer Science Institute VI - Intelligent Systems(自主智能系统,计算机科学研究所VI - 智能系统) ; Robotics, Center for Robotics(机器人学,机器人中心) ; the Lamarr Institute for Machine Learning(拉马尔机器学习研究所) ; Artificial Intelligence, University of Bonn, Germany(人工智能,波恩大学,德国)
专题命中 仿真与规划 :world model(abstract);world model(abstract);model-based reinforcement learning(abstract);分类 cs.AI、cs.LG、cs.RO
AI总结 提出Dream-MPC方法,通过从展开策略生成少量候选轨迹,并利用学习的世界模型进行梯度上升优化,结合不确定性正则化和时间上的优化迭代摊销,显著提升了底层策略性能,在24个连续控制任务上优于无梯度MPC和现有基线。
Comments Accepted for International Conference on Machine Learning (ICML) 2026