予測された未来だけでは不十分:ロボットマニピュレーションのための実行可能な目標の学習
Predicted Futures Are Not Enough: Learning Executable Goals for Robot Manipulation
3Dトレース世界モデルから実行可能な終端目標を明示的に出力する学習インターフェースを提案し、5つのマニピュレーションタスクで平均79.69%の成功率を達成した。
詳しい要約
1. どんなもの?
2. 先行研究と比べてどこがすごい?
3. 技術・手法の肝は?
4. どうやって有効だと検証した?
5. 議論はある?
6. 次に読むべき論文は?
※ AIが要旨から生成した要約です。正確性は原文をご確認ください。
著者: Tzu-Yu Chuang, Ching-Hsiang Chang, Yi-Hsiu Lee, Yi-Ting Chen, Min Sun, YuanFu Yang
分類: cs.RO
原文アブストラクト
Generative world models provide rich predictions of how manipulation scenes may evolve toward task objectives, yet those futures do not directly expose the compact task variables required by control. When training supervises future prediction alone, terminal goal accuracy is not an explicit learning objective, even when geometric recovery is available. We present Entity-Level Goal Readout, a learned prediction-to-execution interface that makes the executable terminal goal an explicit output of a 3D trace world model. It combines object-centric pose prediction with translation grounded in observed depth to produce a compact goal in SE(3). A shared Pose-Native Executor consumes this fixed goal with online object-pose feedback for closed-loop control without rerunning the world model. Across five manipulation tasks, the pipeline achieves a mean success rate of 79.69%. Goal diagnostics directly measure terminal goal accuracy, while controlled translation perturbations characterize how execution degrades under goal error. Zero-shot deployment on a Franka arm achieves 73.33% success on nominal StackCube, 66.67% with distractors, and 75.00% on PickPlate with a target unseen during policy training. These results support treating the prediction-to-execution interface as an explicit learned component of world-model planning rather than incidental post-processing in the control pipeline itself. Project page: https://claire0730.github.io/executable-goals/
関連論文
- 巧みな把持安定性のための時間的視触覚学習マニピュレーション
- FoldBack: 長期的な衣類折り畳みのための自己修正型マスク生成ポリシーマニピュレーション
- 関節特化型ハイブリッド遠隔駆動を用いた全駆動4自由度ロボット指の設計マニピュレーション
- UMI型ロボット教示のための高精度エンドエフェクタ位置推定マニピュレーション
- RoboPace: 接触を考慮した行動チャンク方策の時間最適リタイミングマニピュレーション
- 断続的な視覚喪失に頑健な実ロボットマニピュレーションのための標的モダリティドロップアウトマニピュレーション