MobileVISTA: 移動マニピュレーションにおける姿勢汎化のための生成データ拡張
MobileVISTA: Generative Data Augmentation for Pose Generalization in Mobile Manipulation
単一姿勢の実演データから、視覚観測の拡張と動作の再ターゲティングを組み合わせて姿勢摂動に頑健な訓練データを生成し、ヒューマノイドや双腕ロボットの移動マニピュレーションの汎化性能を向上させるフレームワーク。
詳しい要約
1. どんなもの?
2. 先行研究と比べてどこがすごい?
3. 技術・手法の肝は?
4. どうやって有効だと検証した?
5. 議論はある?
6. 次に読むべき論文は?
※ AIが要旨から生成した要約です。正確性は原文をご確認ください。
著者: Suzannah Wistreich, Stephen Tian, Isabella Huang, Vitor Campagnolo Guizilini, Sergey Zakharov, Katherine Liu, Jiajun Wu
分類: cs.RO, cs.CV, cs.LG
原文アブストラクト
Mobile manipulators such as humanoid robots are increasingly deployed in dynamic, unstructured environments to perform dexterous manipulation tasks. However, end-to-end manipulation policies trained to imitate demonstration data collected from a single robot pose are brittle: even centimeter-scale deviations in robot pose at deployment can drive ego-centric observations and end-effector trajectories out of the training distribution, leading to sharp drops in performance. We introduce MobileVISTA, a data generation framework that transforms demonstrations captured at canonical poses into diverse, pose-perturbed training data by jointly (1) augmenting egocentric visual observations and (2) retargeting actions to compensate for base pose changes. Unlike prior methods, which assume a camera rigidly mounted off the actuated chain or non-trivial articulated robot geometry largely out of frame, MobileVISTA targets compatibility with egocentric platforms (e.g., humanoids) where the camera is both influenced by and must observe the robot's kinematic chain as it moves. We study MobileVISTA in simulated tasks spanning humanoid and bimanual embodiments, and on a real Galaxea R1 Pro. We find policies trained on MobileVISTA-augmented data demonstrate improved robustness to previously out-of-distribution poses encountered at test time, without additional demonstration collection or a trained generative model. Additionally, we find MobileVISTA's benefit is largest on tested humanoids, where the camera rides the actuated chain and the robot fills much of the frame. Additional videos and appendix can be found on our website: https://mobilevista.github.io
関連論文
- RMMBench:ロボット移動マニピュレーションのための包括的ベンチマーク移動マニピュレーション
- ALFRED: 長期植物モニタリングのための要件駆動型オープンソース移動マニピュレータ開発移動マニピュレーション
- LQRとArUcoの融合:二輪ロボットにおけるナビゲーションと非対称マニピュレーションのための堅牢な階層制御移動マニピュレーション
- Zephyron:太陽光支援移動マニピュレータの統合設計と解析的評価によるマルチモーダル環境偵察と分散視覚推論移動マニピュレーション
- 地図認識型視覚運動ポリシーによる移動マニピュレーション移動マニピュレーション
- MoPA: サブシステム特化型知覚アラインメントによる協調的移動マニピュレーション移動マニピュレーション