Jacky Liang
収録論文 23本 ・ フィジカルAI/ロボット学習
模倣学習VLA
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- BPP: 重要な履歴フレームに注目した長文脈ロボット模倣学習模倣学習2026/2/1
視覚言語モデルで検出したタスクに関連するキーフレームのみを条件とすることで、履歴に依存するロボットタスクの模倣学習を安定化させ、実世界タスクで成功率を大幅に向上させた。
- Gemini Robotics 1.5:高度な身体性推論・思考・動作転移で汎用ロボットの最前線を押し広げるVLA2025/10/1
異種ロボットデータから学ぶ動作転移機構と自然言語での多段階推論を組み合わせたVLAモデルと、身体性推論に特化したモデルを発表し、複雑な多段階タスクの実行と解釈性を向上させた。
- Chain-of-Modality: マルチモーダル人間動画と視覚言語モデルによる操作プログラムの学習VLA2025/4/1
筋電アームバンドやマイクなどで計測した人間の操作動画から、視覚言語モデルがタスク計画と力などの制御パラメータを抽出し、ロボットに操作タスクを実行させる手法を提案した。
- Gemini Robotics:AIを物理世界へVLA2025/3/1
Gemini 2.0を基盤に、ロボットを直接制御できる汎用VLAモデル「Gemini Robotics」と、空間・時間理解を強化した推論モデル「Gemini Robotics-ER」を提案した論文。
- 視覚言語モデルは文脈内価値学習器であるVLA2024/11/1
視覚言語モデルにシャッフルした動画フレームの時系列順序を予測させることで、ロボットやタスク固有の訓練なしに300以上の実世界タスクの進捗を推定できる汎用価値関数を実現した。
- PIVOT: Iterative Visual Prompting Elicits Actionable Knowledge for VLMs2024/2/1
- 言語モデル予測制御による人間フィードバックからの高速学習VLA2024/2/1
人間との対話履歴を遷移モデルとしてLLMに微調整し、モデル予測制御と組み合わせることで、ロボットの教示効率と成功率を向上させるフレームワークLMPCを提案。
- Chain of Code: Reasoning with a Language Model-Augmented Code Emulator2023/12/1
- Open X-Embodiment: Robotic Learning Datasets and RT-X Models2023/10/1
- Code as Policies: Language Model Programs for Embodied Control2022/9/1
- Inner Monologue: Embodied Reasoning through Planning with Language Models2022/7/1
- Learning Preconditions of Hybrid Force-Velocity Controllers for Contact-Rich Manipulation2022/6/1
- Search-Based Task Planning with Learned Skill Effect Models for Lifelong Robotic Manipulation2021/9/1
- Visual Identification of Articulated Object Parts2020/12/1
- Contact Localization for Robot Arms in Motion without Torque Sensing2020/11/1
- A Modular Robotic Arm Control Stack for Research: Franka-Interface and FrankaPy2020/11/1
- Learning to Compose Hierarchical Object-Centric Controllers for Robotic Manipulation2020/11/1
- Learning Active Task-Oriented Exploration Policies for Bridging the Sim-to-Real Gap2020/6/1
- In-Hand Object Pose Tracking via Contact Feedback and GPU-Accelerated Robotic Simulation2020/2/1
- DexPilot: Vision Based Teleoperation of Dexterous Robotic Hand-Arm System2019/10/1
- Towards Precise Robotic Grasping by Probabilistic Post-grasp Displacement Estimation2019/9/1
- GPU-Accelerated Robotic Simulation for Distributed Reinforcement Learning2018/10/1
- Dex-Net 2.0: Deep Learning to Plan Robust Grasps with Synthetic Point Clouds and Analytic Grasp Metrics2017/3/1