Playing 20 Question Game with Policy-Based Reinforcement Learning
基于策略的强化学习玩20个问题游戏
机构 * Peking University(北京大学) ; Microsoft Corporation(微软公司) ; Microsoft Development Co., Ltd(微软开发有限公司)
专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.AI
AI总结 提出基于策略的强化学习方法,使提问者通过与用户持续交互学习最优问题选择策略,并利用奖励网络估计信息奖励,无需知识库且对噪声答案鲁棒。
Comments Withdrawal from the conference