DOP: Deep Optimistic Planning with Approximate Value Function Evaluation
专题命中 仿真与规划 :model-based reinforcement learning(abstract);分类 cs.LG、cs.RO
Comments to appear as an extended abstract paper in the Proc. of the 17th International Conference on Autonomous Agents and Multiagent Systems (AAMAS 2018), Stockholm, Sweden, July 10-15, 2018, IFAAMAS. arXiv admin note: text overlap with arXiv:1803.00297