AsyncWebRL: Efficient Asynchronous Reinforcement Learning for Multi-Step Visual Web Agents
AsyncWebRL: 面向视觉网页智能体的高效多步强化学习
机构 * UIUC(伊利诺伊大学香槟分校) ; Microsoft(微软) ; CMU(卡内基梅隆大学)
AI总结 提出异步系统设计和算法改进,解决多步强化学习中GPU空闲和轨迹过长问题,实现训练吞吐量提升2.9倍,并在WebGym测试集上取得新最优结果。
Comments Updated logo and code link