FutureWorld: A Live Reinforcement Learning Environment for Predictive Agents with Real-World Outcome Rewards
FutureWorld: 一个用于预测代理的实时强化学习环境,具有现实世界结果奖励
机构 * College of Software, Nankai University(南开大学软件学院) ; Academy of Mathematics and Systems Science, Chinese Academy of Sciences(中国科学院数学与系统科学研究院) ; School of Computer Science and Technology, University of Science and Technology of China(中国科学技术大学计算机科学与技术学院) ; Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) ; IIIS, Tsinghua University(清华大学智能系统与信息工程研究院) ; Zhongguancun Academy, Beijing, China(北京中关村学院)
AI总结 本文提出FutureWorld,一个实时强化学习环境,通过闭环预测、结果实现与参数更新,提升预测准确性与校准能力。
Comments The code will be released in the near future. The experiments are currently ongoing