MPPI制御のための残差ダイナミクスのタスク指向能動学習
Task-Oriented Active Learning of Residual Dynamics for Model Predictive Path Integral Control
モデル予測パス積分制御(MPPI)において、オンラインのガウス過程残差学習をタスク達成に寄与する情報に基づいて能動的に行う基準ToIAを提案し、オフロードナビゲーションの成功率を向上させた。
詳しい要約
1. どんなもの?
2. 先行研究と比べてどこがすごい?
3. 技術・手法の肝は?
4. どうやって有効だと検証した?
5. 議論はある?
6. 次に読むべき論文は?
※ AIが要旨から生成した要約です。正確性は原文をご確認ください。
著者: Nobuaki Aoki, Hojin Lee, Stefan Sosnowski, Sandra Hirche
分類: cs.RO, eess.SY
原文アブストラクト
Online residual learning can reduce model mismatch in predictive control, but passive data collection may fail to adequately cover states that become important later in the task. Task-agnostic active learning targets uncertain or informative regions, but information acquired in such regions does not necessarily improve task performance. This paper introduces Task-Oriented Information Acquisition (ToIA), an active-learning criterion for model predictive path integral control (MPPI) with online Gaussian process (GP) residual learning. For each sampled control sequence, ToIA estimates how much an observation obtained early in the rollout would reduce predictive uncertainty at later states on the same rollout, and weights this reduction by the rollout's relevance to the task. The score is evaluated over the existing MPPI rollout batch without sampling future observations or re-optimizing control under hypothetical posterior updates. In simulated off-road navigation across held-out maps with heterogeneous terrain, ToIA improved the goal-reaching success rate over passive GP learning by 19.3 and 27.4 percentage points and outperformed task-agnostic active-learning baselines across dense and sparse online-learning intervals. An ablation study indicates that task relevance is particularly important under sparse model updates. The implementation supports online control at 20 Hz on an NVIDIA RTX 2080 Ti.
関連論文
- チャンク型VLAマニピュレーションポリシーの学習と実機展開のためのSim-to-Real統合パイプラインsim2real
- 運動学を超えて:筋駆動模倣学習のためのシミュレーション忠実度ベンチマークsim2real
- CRISP: 多様な形状と接触ソルバを備えた接触リッチロボットシミュレーション基盤sim2real
- 同じ世界、異なる知識:孤立評価が世界モデルの修復を誤判定するときsim2real
- DEXTERA: 単一画像から実機展開可能な巧みなマニピュレーションへ向けたReal-to-Sim-to-Realsim2real
- 単一スキャンからのガウシアンスプラッティングによる実演合成と視覚運動ポリシー学習sim2real