日本フィジカルAI新聞

世界のフィジカルAIを、日本語で。

週刊ニュースレター購読
arXiv:2103.09189

Goal-constrained Sparse Reinforcement Learning for End-to-End Driving

Goal-constrained Sparse Reinforcement Learning for End-to-End Driving

シェア:XThreadsFacebookLINEはてブBluesky

著者: Pranav Agarwal, Pierre de Beaucorps, Raoul de Charette

分類: cs.RO, cs.AI, cs.CV

原文アブストラクト

Deep reinforcement Learning for end-to-end driving is limited by the need of complex reward engineering. Sparse rewards can circumvent this challenge but suffers from long training time and leads to sub-optimal policy. In this work, we explore full-control driving with only goal-constrained sparse reward and propose a curriculum learning approach for end-to-end driving using only navigation view maps that benefit from small virtual-to-real domain gap. To address the complexity of multiple driving policies, we learn concurrent individual policies selected at inference by a navigation system. We demonstrate the ability of our proposal to generalize on unseen road layout, and to drive significantly longer than in the training.