Thomas Kollar
収録論文 19本 ・ フィジカルAI/ロボット学習
VLAマニピュレーション
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- GHIL-Glue: フィルタリングされたサブゴール画像による階層制御VLA2024/10/1
生成モデルが作るサブゴール画像をフィルタリングし、下位方策と効果的に繋ぐことで、言語条件付きロボット操作の汎化性能を向上させた研究。
- 言語埋め込みガウシアンスプラット(LEGS):移動ロボットによる部屋規模表現の漸進的構築VLA2024/9/1
移動ロボットが環境を移動しながら、見た目と言語意味を統合した3D表現(LEGS)をオンラインで構築し、オープン語彙の物体検索を可能にする手法を提案。LERFより3.5倍高速で、最大66%の精度で物体を特定できる。
- OpenVLA: オープンソースの視覚-言語-行動モデルVLA2024/6/1
97万件の実機ロボット実演で学習した7BパラメータのオープンソースVLAモデルを提案し、閉源モデルRT-2-Xを上回る汎用マニピュレーション性能と効率的なファインチューニングを実現した。
- 行動クローニング方策の汎化性能を統計的に評価する枠組みVLA2024/5/1
少数のロールアウトから、行動クローニング方策の性能分布の下限をユーザー指定の信頼度で保証する統計的評価フレームワークを提案し、シミュレーションと実機で検証した。
- DROID: 大規模実環境ロボットマニピュレーションデータセットマニピュレーション2024/3/1
北米・アジア・欧州の50名が12ヶ月かけて564シーン・84タスクで収集した76k軌道・350時間のロボット操作データセットを構築し、学習ポリシーの性能と汎化性能が向上することを示した。
- The Teenager's Problem: Efficient Garment Decluttering as Probabilistic Set Cover2023/10/1
- Open X-Embodiment: Robotic Learning Datasets and RT-X Models2023/10/1
- Bagging by Learning to Singulate Layers Using Interactive Perception2023/3/1
- CARTO: Category and Joint Agnostic Reconstruction of ARTiculated Objects2023/3/1
- HANDLOOM: Learned Tracing of One-Dimensional Objects for Inspection and Manipulation2023/3/1
- Language-Driven Representation Learning for Robotics2023/2/1
- AutoBag: Learning to Open Plastic Bags and Insert Objects2022/10/1
- SGTM 2.0: Autonomously Untangling Long Cables using Interactive Perception2022/9/1
- ShAPO: Implicit Representations for Multi-Object Shape, Appearance, and Pose Optimization2022/7/1
- Efficiently Learning Single-Arm Fling Motions to Smooth Garments2022/6/1
- CenterSnap: Single-Shot Multi-Object 3D Shape Reconstruction and Categorical 6D Pose and Size Estimation2022/3/1
- SimNet: Enabling Robust Unknown Object Manipulation from Pure Synthetic Data via Stereo2021/6/1
- A Mobile Manipulation System for One-Shot Teaching of Complex Tasks in Homes2019/10/1
- Generalized Grounding Graphs: A Probabilistic Framework for Understanding Grounded Commands2017/12/1