自動運転車の安全クリティカル制御のための強化学習
Reinforcement learning for safety-critical control of an automated vehicle
強化学習のPPOを用いて移動ロボットSPIDERを障害物回避しながら目標地点へ誘導する制御器を訓練し、シミュレーションと実機で検証した。
著者: Florian Thaler, Franz Rammerstorfer, Jon Ander Gomez, Raul Garcia Crespo, Leticia Pasqual, Markus Postl
分類: cs.RO
原文アブストラクト
We present our approach for the development, validation and deployment of a data-driven decision-making function for the automated control of a vehicle. The decisionmaking function, based on an artificial neural network is trained to steer the mobile robot SPIDER towards a predefined, static path to a target point while avoiding collisions with obstacles along the path. The training is conducted by means of proximal policy optimisation (PPO), a state of the art algorithm from the field of reinforcement learning. The resulting controller is validated using KPIs quantifying its capability to follow a given path and its reactivity on perceived obstacles along the path. The corresponding tests are carried out in the training environment. Additionally, the tests shall be performed as well in the robotics situation Gazebo and in real world scenarios. For the latter the controller is deployed on a FPGA-based development platform, the FRACTAL platform, and integrated into the SPIDER software stack.
関連論文
- 動的障害物環境での安全飛行のための予測リスク誘導強化学習強化学習/ナビゲーション
- ManeuverNet: ダブルアクカーマン操舵ロボットの精密操縦のための報酬関数最適化を伴うSoft Actor-Criticフレームワーク強化学習/ナビゲーション
- 二重アクカーマン操舵ロボットの安全な機動のためのSoft Actor-Criticフレームワーク強化学習/ナビゲーション
- 分布強化学習に基づく自律水上艇の統合的意思決定と制御強化学習/ナビゲーション
- 渦流を横切るロボット魚の効率的ナビゲーション強化学習/ナビゲーション