Andy Zeng
収録論文 45本 ・ フィジカルAI/ロボット学習
VLAマニピュレーション基盤モデル応用
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- ロボットナビゲーションと操作のためのマルチモーダル空間言語マップVLA2025/6/1
視覚・言語・音声の特徴を3D環境再構成に融合した空間マップを構築し、LLMと組み合わせて自然言語指示やマルチモーダルな目標を地図上に定位してロボットのナビゲーションと操作を可能にする手法を提案。
- 二本腕の身体性AI:ゼロショット学習、安全性、モジュール性マニピュレーション2024/4/1
自然言語指示に基づき二本腕ロボットが協調して長期的タスクをゼロショットで実行するモジュール型身体性AIシステムを提案。安全性とモジュール性を重視し、実世界データなしで動作する。
- 基盤モデルの実世界ロボット応用:レビュー基盤モデル応用2024/2/1
LLMやVLMなどの基盤モデルを実ロボットシステムの知覚・運動計画・制御にどう組み込むかを、入出力関係の観点から整理したレビュー。
- 言語モデル予測制御による人間フィードバックからの高速学習VLA2024/2/1
人間との対話履歴を遷移モデルとしてLLMに微調整し、モデル予測制御と組み合わせることで、ロボットの教示効率と成功率を向上させるフレームワークLMPCを提案。
- PIVOT: Iterative Visual Prompting Elicits Actionable Knowledge for VLMs2024/2/1
- Generative Expressive Robot Behaviors using Large Language Models2024/1/1
- Chain of Code: Reasoning with a Language Model-Augmented Code Emulator2023/12/1
- Distilling and Retrieving Generalizable Knowledge for Robot Manipulation via Language Corrections2023/11/1
- Video Language Planning2023/10/1
- Large Language Models as General Pattern Machines2023/7/1
- Robots That Ask For Help: Uncertainty Alignment for Large Language Model Planners2023/7/1
- Rearrangement Planning for General Part Assembly2023/7/1
- Language to Rewards for Robotic Skill Synthesis2023/6/1
- TidyBot: Personalized Robot Assistance with Large Language Models2023/5/1
- RoboPianist: Dexterous Piano Playing with Deep Reinforcement Learning2023/4/1
- Audio Visual Language Maps for Robot Navigation2023/3/1
- Grounded Decoding: Guiding Text Generation with Grounded Models for Embodied Agents2023/3/1
- PaLM-E: An Embodied Multimodal Language Model2023/3/1
- MIRA: Mental Imagery for Robotic Affordances2022/12/1
- VIRDO++: Real-World, Visuo-tactile Dynamics and Perception of Deformable Objects2022/10/1
- Visual Language Maps for Robot Navigation2022/10/1
- Code as Policies: Language Model Programs for Embodied Control2022/9/1
- Inner Monologue: Embodied Reasoning through Planning with Language Models2022/7/1
- Do As I Can, Not As I Say: Grounding Language in Robotic Affordances2022/4/1
- Learning Pneumatic Non-Prehensile Manipulation with a Mobile Blower2022/4/1
- Learning to Fold Real Garments with One Arm: A Case Study in Cloud-Based Robotics Research2022/4/1
- Implicit Kinematic Policies: Unifying Joint and Cartesian Action Spaces in End-to-End Robot Learning2022/3/1
- Multiscale Sensor Fusion and Continuous Control with Neural CDEs2022/3/1
- VIRDO: Visio-tactile Implicit Representations of Deformable Objects2022/2/1
- Multi-Task Learning with Sequence-Conditioned Transporter Networks2021/9/1
- Implicit Behavioral Cloning2021/9/1
- Learning to See before Learning to Act: Visual Pre-training for Manipulation2021/7/1
- XIRL: Cross-embodiment Inverse Reinforcement Learning2021/6/1
- Spatial Intention Maps for Multi-Agent Mobile Manipulation2021/3/1
- Learning to Rearrange Deformable Cables, Fabrics, and Bags with Goal-Conditioned Transporter Networks2020/12/1
- Transporter Networks: Rearranging the Visual World for Robotic Manipulation2020/10/1
- Spatial Action Maps for Mobile Manipulation2020/4/1
- Grasping in the Wild:Learning 6DoF Closed-Loop Grasping from Low-Cost Demonstrations2019/12/1
- ClearGrasp: 3D Shape Estimation of Transparent Objects for Manipulation2019/10/1
- Form2Fit: Learning Shape Priors for Generalizable Assembly from Disassembly2019/10/1
- DensePhysNet: Learning Dense Physical Object Representations via Multi-step Dynamic Interactions2019/6/1
- TossingBot: Learning to Throw Arbitrary Objects with Residual Physics2019/3/1
- Learning Synergies between Pushing and Grasping with Self-supervised Deep Reinforcement Learning2018/3/1
- Robotic Pick-and-Place of Novel Objects in Clutter with Multi-Affordance Grasping and Cross-Domain Image Matching2017/10/1
- Multi-view Self-supervised Deep Learning for 6D Pose Estimation in the Amazon Picking Challenge2016/9/1