Reinforced sequential Monte Carlo for amortised sampling
强化序贯蒙特卡洛用于摊销采样
机构 * University of Edinburgh ; Mila -- Qu\'ebec AI Institute ; CIFAR Fellow
专题命中 其他多模态 :multi-modal(abstract)
AI总结 本文提出一种摊销方法与粒子方法相结合的采样框架,通过最大熵强化学习训练序贯蒙特卡洛采样器,并利用离线策略学习提高目标分布探索效率,在合成多模态目标和丙氨酸二肽构象玻尔兹曼分布上验证了改进的近似精度与训练稳定性。
Comments ICML 2026. Code: https://github.com/hyeok9855/ReinforcedSMC