全身UMI:実時間動作生成によるヒューマノイド全身マニピュレーションへのUMIスキル転移
Whole-Body UMI: Transferring UMI Manipulation Skills to Humanoid Whole-Body Manipulation via Real-Time Motion Generation
UMIの手先軌道データからヒューマノイドの全身協調動作を実時間生成し、拡散方策と組み合わせて4タスクで全身マニピュレーションを実現した。
著者: Yuxuan Nai, Leixin Chang, Liangjing Yang, Shuo Yang, Zhongyu Li
分類: cs.RO
原文アブストラクト
Collecting whole-body demonstrations for humanoid manipulation mostly relies on teleoperation, which is costly and hard to scale up. The Universal Manipulation Interface (UMI) provides a scalable data collection paradigm, but end-effector trajectories alone underdetermine humanoid whole-body coordination, which is insufficient for whole-body demonstration collection. Therefore, we introduce Whole-Body UMI (WB-UMI), a task-agnostic, real-time and end-effector conditioned motion generator that decouples whole-body coordination learning from task semantics learning through a shared end-effector interface. A diffusion policy learns from native UMI demonstrations, while WB-UMI learns independently from retargeted motion capture, requiring no body trackers or paired image--whole-body demonstrations during task-specific data collection. In real deployment, an asynchronous hierarchy integrates the diffusion policy, motion generator, and a whole-body controller with latency compensation and measured-state feedback. Real-robot experiments on G1 support real-time closed-loop transfer across four tasks, achieving 90% success in drawer closing, 80% in shelf pick-and-place, 30% in ball toss, and 40% in Loco-PnP, which shows the effectiveness of this hierarchy in transferring native UMI skills to humanoid whole-body manipulation.
関連論文
- LYRIC: 言語駆動の物理ベース全身接触リッチ物体インタラクション制御全身マニピュレーション
- WholeBodyWAM:スケーラブルな動作事前分布を用いた全身ワールドアクションモデルの学習全身マニピュレーション
- InterPrior: 物理ベースの人間-物体インタラクションのための生成的制御のスケーリング全身マニピュレーション
- AdaptManip: オンライン再帰的状態推定による適応的な全身物体持ち上げ・運搬の学習全身マニピュレーション
- HumanoidExo: ウェアラブル外骨格によるスケーラブルな全身ヒューマノイド操作全身マニピュレーション
- 人間型ロボットによる大型物体の抱え込み:強化学習を用いた全身マニピュレーション全身マニピュレーション