深層Qネットワーク強化学習を用いた発電所における無人航空機の自己点検手法
Self-Inspection Method of Unmanned Aerial Vehicles in Power Plants Using Deep Q-Network Reinforcement Learning
発電所の点検作業を自律的に行うため、深層Qネットワーク(DQN)強化学習を用いてUAVのナビゲーションを訓練し、風やバッテリー残量などの環境要因を考慮した報酬設計により実用的な点検戦略を実現した。
著者: Haoran Guan
分類: cs.RO, cs.LG
原文アブストラクト
For the purpose of inspecting power plants, autonomous robots can be built using reinforcement learning techniques. The method replicates the environment and employs a simple reinforcement learning (RL) algorithm. This strategy might be applied in several sectors, including the electricity generation sector. A pre-trained model with perception, planning, and action is suggested by the research. To address optimization problems, such as the Unmanned Aerial Vehicle (UAV) navigation problem, Deep Q-network (DQN), a reinforcement learning-based framework that Deepmind launched in 2015, incorporates both deep learning and Q-learning. To overcome problems with current procedures, the research proposes a power plant inspection system incorporating UAV autonomous navigation and DQN reinforcement learning. These training processes set reward functions with reference to states and consider both internal and external effect factors, which distinguishes them from other reinforcement learning training techniques now in use. The key components of the reinforcement learning segment of the technique, for instance, introduce states such as the simulation of a wind field, the battery charge level of an unmanned aerial vehicle, the height the UAV reached, etc. The trained model makes it more likely that the inspection strategy will be applied in practice by enabling the UAV to move around on its own in difficult environments. The average score of the model converges to 9,000. The trained model allowed the UAV to make the fewest number of rotations necessary to go to the target point.
関連論文
- GNSS劣化環境下の都市部におけるクアッドロータUAVナビゲーションのための信念適応型オンライン自律性UAVナビゲーション
- RACO: 点検指向UAV視覚言語ナビゲーションのための信頼性を考慮した粗目標最適化UAVナビゲーション
- AirAlign: ジオメトリを考慮したUAV最終メートル航法のための相対ポーズ整列UAVナビゲーション
- 形状適合領域による閉鎖・劣化環境での飛行UAVナビゲーション
- PILOT: 部分観測下での自律UAVのエンドツーエンド動作計画のための特権模倣学習UAVナビゲーション
- FlowPilot: 俊敏なUAVナビゲーションのためのリアルタイム世界行動モデリングUAVナビゲーション