バッテリー状態を考慮した強化学習によるアグレッシブなクアッドロータ飛行
Battery-Aware Reinforcement Learning for Aggressive Quadrotor Flight
バッテリー電圧の低下を考慮した強化学習制御により、ドローンレースなどのアグレッシブ飛行で追従誤差とラップタイムを改善した研究。
詳しい要約
1. どんなもの?
2. 先行研究と比べてどこがすごい?
3. 技術・手法の肝は?
4. どうやって有効だと検証した?
5. 議論はある?
6. 次に読むべき論文は?
※ AIが要旨から生成した要約です。正確性は原文をご確認ください。
著者: Alejandro Sanchez Roncero, Olov Andersson, Petter Ogren
分類: cs.RO
原文アブストラクト
Agile flight tasks such as drone racing and pursuit-evasion require strong acceleration and precise turns, but the available thrust changes as the battery discharges and voltage drops under load. Conservative command limits make this variation easier to tolerate, at the cost of unused performance. We investigate how learned controllers can use that additional thrust while retaining the flight controller's voltage compensation and rate control. Our training simulator couples an identified load-transient battery model to rotor dynamics and firmware saturation. The feedforward policy receives filtered voltage during both training and deployment. Controlled ablations distinguish the benefit of a larger thrust-command range from that of voltage information. On a 38 g Crazyflie Brushless, the resulting policy reduces circle tracking error by 49% relative to stock-authority RL at 3.84 m/s, while preserving easy-task precision. Mean 20-lap race time decreases from 106.22 s to 95.24 s. Compared with a voltage-blind policy with the same increased authority, hardware error and race time are lower by 15.3% and 4.5%, respectively. In simulation, replacing the policy's voltage input with a recording from a different battery condition worsens hard-circle tracking, with a smaller, voltage-dependent effect in racing. Together, these results show where a simple voltage input complements existing actuator compensation in aggressive learned flight.
関連論文
- CORB-Planner: 高速飛行におけるRLプランニングのための観測としてのコリドー飛行制御/強化学習
- 強化学習によるクアッドコプターの耐故障制御飛行制御/強化学習
- 大きな外乱を受けるUASのRLに基づく制御飛行制御/強化学習
- スケール対応深層強化学習による動特性不変なクアッドロータ制御飛行制御/強化学習
- 深層強化学習によるソフトアクチュエータ昆虫サイズ飛行ロボットのホバリング飛行飛行制御/強化学習
- 深層強化学習によるマルチロータ航空ロボットの運動制御飛行制御/強化学習