FlashDexRetarget: 多動作リターゲティングによる器用操作データ生成の高速化
FlashDexRetarget: Accelerating Dexterous Manipulation Data Generation through Multi-Motion Retargeting
人間の手と物体のデモンストレーションをロボットハンドに転送する強化学習ベースのリターゲティング手法を提案し、50動作で90%の成功率を達成しつつ、学習計算量を最大100倍削減した。
詳しい要約
1. どんなもの?
2. 先行研究と比べてどこがすごい?
3. 技術・手法の肝は?
4. どうやって有効だと検証した?
5. 議論はある?
6. 次に読むべき論文は?
※ AIが要旨から生成した要約です。正確性は原文をご確認ください。
著者: Kyungmin Lee, Sibeen Kim, Dongyoon Hwang, Yoonsang Oh, Donghu Kim, Youngdo Lee, I Made Aswin Nahrendra, Jaegul Choo, Hojoon Lee
分類: cs.RO
原文アブストラクト
Human hand-object demonstrations offer a reusable source of dexterous robot manipulation data, but transferring them across embodiments requires physically feasible retargeting. Existing physics-based approaches face limitations in retargeting success, motion-specific training efficiency, or both. To address these limitations, we introduce FlashDexRetarget, an RL-based framework for high-success, efficient dexterous motion retargeting. To make the demonstrated interaction easier to learn, we combine object point-cloud observations, hand-object distance features, and future trajectory encodings with complementary rewards that supervise object motion and reference hand-object relationships. To further accelerate learning, we employ separate left- and right-hand actor critic networks and adapt the off-policy algorithm, FlashSAC to dexterous motion tracking. On a benchmark of 50 motions spanning single-object and two-object interactions, FlashDexRetarget achieves a 90% success rate, approximately 2.5x that of the evaluated sampling-based baselines, while requiring up to 100x less training compute than the evaluated RL-based baselines. Evaluations on both XHand and Sharpa Wave Hand show consistent gains, and component-wise ablations examine the contributions of our design choices. Beyond the 50-motion benchmark, experiments with 200, 500, and 1,000 motions demonstrate that our method remains stable at larger scales and produces successful retargeted motions more efficiently as the training set grows. Qualitative replay results using real-world-captured demonstrations further illustrate the applicability of our framework to recorded human manipulation. Videos and code are available at https://davian-robotics.github.io/FlashDexRetarget/
関連論文
- スクリューアテンション:Transformer内部に剛体代数を組み込むマニピュレーション
- SkeleWAM: 骨格ワールドアクションモデリングによる効率的なロボットマニピュレーションマニピュレーション
- 解像度に一貫したヤコビアン場を学習する生体模倣剛柔指マニピュレーション
- Recova: 自律ロボットマニピュレーションのためのエージェント誘導型失敗回復マニピュレーション
- 経験と実演による6自由度把持合成の継続学習マニピュレーション
- 再構成・練習・実世界展開:身体性エージェントのためのガイド付き自己改善マニピュレーション