RoaD: Rollouts as Demonstrations for Closed-Loop Supervised Fine-Tuning of Autonomous Driving Policies
RoaD: 通过回滚作为演示实现自动驾驶策略的闭环监督微调
机构 * NVIDIA Research(NVIDIA研究部) ; Huawei VN Research Center(华为越南研究中心) ; Stanford University(斯坦福大学)
专题命中 仿真评测 :autonomous driving(title,abstract);end-to-end driving(abstract);分类 cs.RO、cs.CV、cs.AI
AI总结 RoaD通过利用自动驾驶策略自身的闭环回滚作为额外训练数据,有效缓解协变量偏移问题,提升闭环监督微调性能。
Comments Preprint