The Geometry of Learning to Avoid Interventions
从紧急停止干预中学习鲁棒性干预
机构 * Paul G. Allen School of Computer Science \& Engineering, University of Washington, Seattle, USA ; Google DeepMind, Seattle, USA
AI总结 本文提出残差干预微调算法,通过结合先验策略解决干预信号不明确的问题,实现鲁棒的策略改进。
高校专区
从紧急停止干预中学习鲁棒性干预
机构 * Paul G. Allen School of Computer Science \& Engineering, University of Washington, Seattle, USA ; Google DeepMind, Seattle, USA
AI总结 本文提出残差干预微调算法,通过结合先验策略解决干预信号不明确的问题,实现鲁棒的策略改进。