RAGrasp: 幾何・意味テンプレート検索と把持転移
RAGrasp: Geometry-Semantic Template Retrieval and Grasp Transfer
展開環境で収集した少数のRGB-D把持テンプレートを検索し、SAM2とDINOv2特徴で対象物を切り出して把持を転移する、再学習不要の平面パラレルジョー把持パイプライン。
詳しい要約
1. どんなもの?
2. 先行研究と比べてどこがすごい?
3. 技術・手法の肝は?
4. どうやって有効だと検証した?
5. 議論はある?
6. 次に読むべき論文は?
※ AIが要旨から生成した要約です。正確性は原文をご確認ください。
著者: Shenzhe Zhu, Chengxiao He, Jan Harder
分類: cs.AI, cs.RO
原文アブストラクト
We present RAGrasp, a retrieval-augmented pipeline for planar parallel-jaw grasping from a compact set of locally collected, grasp-annotated RGB-D (color and depth) templates. Unlike task-specific predictors trained primarily on large public or synthetic grasp datasets, RAGrasp requires no end-to-end retraining for a new deployment.Its template memory is constructed from observations collected with the deployment camera, robot, and gripper in the target workspace, thereby aligning stored examples with the local sensing and embodiment conditions. The system uses self-supervised DINOv2 visual fea- tures together with appearance and depth cues to prompt the Segment Anything Model 2 (SAM2), which isolates the query object. A two-stage geometry-semantic retrieval cascade then selects a template, and a confidence gate chooses one of two grasp- transfer estimators. The transferred grasp is refined using mask- support and silhouette-contact constraints before calibrated 2D- to-3D conversion. In real-world trials, RAGrasp achieves 20/20 successful grasps on seen objects and 19/20 on unseen objects. Within the evaluated setting, the results demonstrate deployment- specific grasp adaptation from limited local annotation and tolerance to the tested viewpoint and illumination changes.
関連論文
- 不確実環境下での柔軟なリーチングのための仮想モデル制御マニピュレーション
- オープンエンド環境におけるロバストな把持マニピュレーションに向けてマニピュレーション
- DexForge: 高忠実度な物理情報に基づく巧みなリターゲティングマニピュレーション
- STC-MPM:軟組織切断における変形・損傷進展・切開形成の連成マニピュレーション
- デモンストレーションで調整するポート・ハミルトン型マニピュレーション方策の再チューニングマニピュレーション
- 能動推論制御のための生成軌道モデルのベンチマークマニピュレーション