能動推論制御のための生成軌道モデルのベンチマーク
Benchmarking Generative Trajectory Models for Active-Inference Control
生成軌道モデルを用いた能動推論制御(GenAIF)を提案し、拡散モデル・自己回帰Transformer・CVAE・フローマッチングをMuJoCoマニピュレーションで比較評価した。
著者: Yulin Li, Mohsen A. Jafari, Andrea Matta
分類: cs.RO
原文アブストラクト
Learning from trajectory demonstrations offers a route to active-inference control of complex systems whose dynamics are difficult to model explicitly. We introduce generative active-inference control (GenAIF), in which one generative trajectory model learns from demonstrations and measured action interventions to supply a goal-conditioned policy distribution and a state-to-observation likelihood mapping. From this control design, we derive three model requirements: (i) useful action proposals, (ii) accurate prediction under imposed actions, and (iii) probabilistic observation evidence for belief updating and expected information gain. We benchmark diffusion, autoregressive Transformers, conditional variational autoencoders (CVAEs), and flow matching in a MuJoCo manipulation task with multiple physical conditions. Diffusion delivers the strongest control across the tested dynamics, while CVAE combines comparable short-horizon prediction with much faster inference. Correct conditioning is decisive, and trajectory reuse offers further computational savings. With the same frozen models, a hidden-dynamics experiment demonstrates prompt belief adaptation after an unannounced tilt change; subsequent instability identifies sustained inference as a remaining challenge. These findings support the use of shared generative trajectory models to connect action proposal, controlled prediction, and observation evidence within GenAIF.
関連論文
- 不確実環境下での柔軟なリーチングのための仮想モデル制御マニピュレーション
- オープンエンド環境におけるロバストな把持マニピュレーションに向けてマニピュレーション
- DexForge: 高忠実度な物理情報に基づく巧みなリターゲティングマニピュレーション
- STC-MPM:軟組織切断における変形・損傷進展・切開形成の連成マニピュレーション
- デモンストレーションで調整するポート・ハミルトン型マニピュレーション方策の再チューニングマニピュレーション
- 物理残差ダイナミクスと低次元全身計画による障害物回避型人間-ロボット協調布運搬マニピュレーション