FlowR2A: Learning Reward-to-Action Distribution for Multimodal Driving Planning
FlowR2A: 学习奖励到动作分布用于多模态驾驶规划
机构 * The University of Hong Kong(香港大学) ; Changan Automobile(长安汽车)
专题命中 多模态Agent :multimodal(title,abstract);分类 cs.AI
AI总结 提出FlowR2A,通过流匹配解码器学习奖励条件动作分布,统一了基于评分和基于锚点的方法,实现密集监督与动态生成,在NAVSIM基准上取得最优结果。
Comments Project page: https://lixirui142.github.io/flowr2a-ad