Hierarchical Reinforcement Learning with Optimal Level Synchronization Based on Flow-Based Deep Generative Model
基于流基深度生成模型的最优层级同步分层强化学习
AI总结 研究强化学习中高维状态等复杂场景,提出基于流基深度生成模型支持直接离策略校正的新型分层强化学习模型,利用模型逆操作与解决其局限性,经实验验证该模型性能优于现有模型。
Comments Published in the Journal of Artificial Intelligence Research (JAIR), Volume 86, 2026. This version corresponds to the published JAIR article
Journal ref Journal of Artificial Intelligence Research, 86, Article 28 (2026), 25 pages