DynaForge: 動的マニピュレーションのデモ生成のための計画誘導型残差学習
DynaForge: Planning-Guided Residual Learning for Dynamic Manipulation Demonstration Generation
低周波の大域計画と高周波の物体中心逆運動学を組み合わせ、残差方策で接触時の動的相互作用を補正することで、動的物体操作の高品質なデモンストレーションを生成するフレームワーク。
詳しい要約
1. どんなもの?
2. 先行研究と比べてどこがすごい?
3. 技術・手法の肝は?
4. どうやって有効だと検証した?
5. 議論はある?
6. 次に読むべき論文は?
※ AIが要旨から生成した要約です。正確性は原文をご確認ください。
著者: Yiyang Jin, Yu Zheng, Xiao He, Hesheng Wang
分類: cs.RO
原文アブストラクト
Dynamic object manipulation is essential for robots operating in real-world environments, yet methods for generating high-quality demonstrations remain limited. Methods designed for static tasks do not readily transfer to dynamic settings. Among dynamic demonstration generators, planning-based methods can fail near contact, while DOMINO-style replay simplifies dynamic interactions and may limit the experience available for policy learning. We present DynaForge, a planning-guided framework that learns residual corrections for dynamic manipulation demonstration generation. DynaForge combines low-frequency global planning with high-frequency object-centric inverse kinematics across task phases, and applies a residual policy to correct actions during dynamic interaction. An implicit curriculum groups rollouts under matched conditions and selects mixed-success groups, focusing residual reinforcement learning on the evolving competence frontier. On Can and Bottle, it uses 0.73x as many optimizer steps as vanilla GRPO at the same nominal environment-step budget, with higher observed final success rates. Across nine simulation tasks, DynaForge increases mean demonstration-generation success from 41.30% of the planning prior to 78.37%. With 800 demonstrations per task, DP3 policies trained on DynaForge data achieve 49.11% mean success, compared with 7.07% for DOMINO data. On three real-world dynamic tasks, DynaForge-trained policies achieve 30-60% success, compared with 0-10% for DOMINO-trained policies, showing the ability of DynaForge for sim-to-real transfer.
関連論文
- エンドタスク成功を超えて:ロボティクスにおける視覚経験検索の監査手法マニピュレーション
- 複数把持点における非無視可能な物理応答を伴う線形変形物体の安定性保証付きマニピュレーションマニピュレーション
- PAKT: 強化学習のための物理的整合性を備えたキネステティック教示マニピュレーション
- 不完全データを活用した高精度ロボットマニピュレーションマニピュレーション
- SafeLoop: 視覚言語行動マニピュレーションのためのリスク認識ロールバックマニピュレーション
- ローカルコーディングエージェントによるマニピュレーションスキルの汎化マニピュレーション