油圧クレーンによる丸太山積み処理のための点群からの把持位置学習
Learning Grasp Targeting from Point Clouds for Log Pile Clearing on a Hydraulic Crane
未分割点群からグラップルの把持位置と向きを学習し、油圧式林業クレーンで丸太山を片付ける模倣学習と強化学習の手法を提案。実機実験で幾何学的ヒューリスティックより多くの丸太を運搬できることを示した。
詳しい要約
1. どんなもの?
2. 先行研究と比べてどこがすごい?
3. 技術・手法の肝は?
4. どうやって有効だと検証した?
5. 議論はある?
6. 次に読むべき論文は?
※ AIが要旨から生成した要約です。正確性は原文をご確認ください。
著者: George Sideris, Lucas Bessai, Heshan Fernando, Elie Ayoub, Nicolas Lemieux, Inna Sharf
分類: cs.RO, cs.LG
原文アブストラクト
In mill yards, log loaders clear dense piles by a sequence of bundle grasps: hundreds of logs rest in contact, and each removal changes the pile available to the next grasp. A learned policy chooses where to place and orient the grapple from unsegmented point clouds and runs on a trailer-mounted hydraulic forestry crane. The policy classifies at which observed point to grasp and predicts depth and grapple orientation there. The same network outputs support behavior cloning (BC), reinforcement learning (RL), and deployment. BC learns from successful top-of-pile demonstrations; RL explores for improvements by fine-tuning the cloned policy (BC$\to$RL) or by training from scratch. In simulation, BC clears 98 of 100 piles of 200 logs, while BC$\to$RL improves load stability. Twelve field trials compare a geometric heuristic, RL from scratch, BC, and BC$\to$RL through complete grasp-transport-deposit cycles. BC and BC$\to$RL deposit 93.8% and 88.9% of pooled inventory, against 80.4% for the heuristic. BC$\to$RL deposits logs on 83.6% of its cycles, against 79.6% for the heuristic and 65.7% for BC, while its simulated stability gain does not carry over to the crane testbed. Trained entirely in simulation and run unchanged on the crane, the learned policies clear more than the hand-filtered heuristic while observing unfiltered clouds that still contain the storage rack's rails and poles.
関連論文
- 密度関数を用いた安全なマルチロボット協調搬送マニピュレーション
- コンパクトなロボットポリシーに必要なのは細粒度の視覚表現マニピュレーション
- ずれた座標系を見抜く:視覚・力覚精密組立のための特権ノイズ蒸留マニピュレーション
- 把持後における物体再配向によるロボット挿入の運動学的修復マニピュレーション
- 未来が一致する間だけコミット:ロボットマニピュレーションのための結果認識型適応アクションチャンキングマニピュレーション
- TacZero: 触覚フィードバックと汎用視覚言語モデルによる訓練不要のペグ挿入マニピュレーション