ReDex: 指先の柔軟なインタラクションによるSim-to-Real巧みな操作ポリシーの修復
ReDex: Repairing Sim-to-Real Dexterous Policies by Finger-Level Compliant Interaction
シミュレーションで訓練した多指ハンドポリシーに対し、実機で人間が指先の接触失敗のみを修正し、その際の力覚情報を模倣学習で取り込むことで、触覚シミュレーションなしに実世界へ適応させる手法。
詳しい要約
1. どんなもの?
2. 先行研究と比べてどこがすごい?
3. 技術・手法の肝は?
4. どうやって有効だと検証した?
5. 議論はある?
6. 次に読むべき論文は?
※ AIが要旨から生成した要約です。正確性は原文をご確認ください。
著者: Jinzhou Li, Hadi Tabatabaee, Kelin Yu, Yuyin Sun, Cheng-Hao Kuo, Roberto Martín-Martín, Nima Fazeli, X. Alice Wu, Xianyi Cheng
分類: cs.RO
原文アブストラクト
Dexterous manipulation policies trained in simulation often fail to transfer to the real world because of errors in contact timing and force regulation. Yet these policies can retain useful multi-finger coordination for task progression. We propose ReDex, a framework for adapting a simulation-trained base policy to the real world by correcting local contact failures and incorporating tactile feedback. Starting from a proprioception-only base policy, ReDex allows a human operator to physically correct contact failures at selected fingers under compliant control during real-world rollouts, while the frozen base policy continues to control the remaining fingers. These rollouts combine base policy execution, human-corrected finger motion, and fingertip force observations. We reconstruct force-informed targets from these rollouts to train a standalone force-conditioned policy via behavior cloning. This design reduces human correction effort, enables learning of contact regulation from real-world interaction, and introduces force feedback into a proprioception-only policy without tactile simulation or complex full-hand teleoperation. We evaluate ReDex on two challenging, contact-rich dexterous manipulation tasks on real hardware. Compared with sim-to-real transferred base policies, ReDex increases Object Flipping success rate from 14\% to 86\% across two objects and average Screwdriver Rotation progress from 26.0% to 95.3% across three objects.
関連論文
- EmbodiedSmith:シミュレーションにおける再帰的自己改善フライホイールによる身体性データのスケーリングsim2real
- 安全なリアルタイムロボット制御のためのマイクロニューラルポリシーsim2real
- 脚から車輪へ:移動ベースヒューマノイドのための身体性を考慮した人間動作リターゲティングsim2real
- ロボットはその記述ではない:形態認識ポリシーの表現堅牢性を評価するGaugeBenchsim2real
- SMART: 大規模合成事前学習によるゼロショットSim-to-Real関節物体マニピュレーションsim2real
- UWB×Crazyflow:空中ロボティクス向け劣化フィードバックの大規模シミュレーションsim2real