Chris Paxton
収録論文 60本 ・ フィジカルAI/ロボット学習
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- 能動的知覚を備えたロボットの計画と状況対応タスク計画/能動的知覚2026/4/1
ロボットの実行中に発生する予期せぬ状況を能動的に知覚し対処するためのフレームワークVAP-TAMPを提案した。視覚言語モデルとシーングラフを活用し、タスクと動作の統合計画を行う。
- LLM-GROP: 大規模言語モデルによる視覚に基づくロボットのタスク・動作計画タスク・動作計画2025/11/1
大規模言語モデルの常識知識と視覚情報を活用し、複数物体の移動を伴うモバイルマニピュレーションのタスク・動作計画を統合するフレームワークを提案。実環境とシミュレーションで成功率と効率を評価した。
- ニューラル運動計画のためのカスケード拡散モデル運動計画2025/5/1
拡散ポリシーを用いて複雑な環境での衝突のない大域的な運動計画を学習するカスケード階層モデルを提案し、ナビゲーションやマニピュレーションで既存手法を約5%上回る性能を示した。
- GraphEQA: 3Dセマンティックシーングラフを用いたリアルタイム身体性質問応答VLA2024/12/1
未知環境での身体性質問応答のため、リアルタイム3Dメトリックセマンティックシーングラフとタスク関連画像をマルチモーダルメモリとしてVLMに接地し、階層的計画で探索・回答する手法を提案。シミュレーションと実環境で成功率向上と計画ステップ削減を示した。
- DynaMem: オープンワールド移動マニピュレーションのためのオンライン動的時空間セマンティックメモリ移動マニピュレーション2024/11/1
ロボットが環境の変化に合わせて3Dメモリを動的に更新し、移動・出現・消滅する物体を扱えるようにする手法を提案。実機で非静止物体のピック・アンド・ドロップ成功率70%を達成した。
- ロボットユーティリティモデル:新環境へのゼロショット展開のための汎用ポリシーマニピュレーション2024/9/1
ファインチューニングなしで新しい環境や未知の物体に直接適応できるゼロショットロボットポリシーを訓練・展開するフレームワークを提案し、5つの実用的タスクで平均90%の成功率を達成した。
- DegustaBot: 個人の好みに基づくマルチオブジェクト再配置のためのゼロショット視覚嗜好推定マニピュレーション2024/7/1
視覚言語基盤モデルとゼロショット視覚プロンプトを用いて、家庭内の物体再配置タスクを個人の視覚的嗜好に合わせて解く手法を提案し、シミュレーションで評価した。
- HACMan++: 空間的に接地された運動プリミティブによるマニピュレーションマニピュレーション2024/7/1
把持や押しなどの運動プリミティブを「種類・接地位置・実行パラメータ」で表現し、強化学習で組み合わせて長期的な操作タスクを達成する手法を提案。形状や姿勢の変化に汎化し、sim-to-real転移も実現した。
- 家庭内でのオープンワールド移動マニピュレーションに向けて:NeurIPS 2023 HomeRobotオープン語彙移動マニピュレーションチャレンジからの教訓移動マニピュレーション2024/7/1
家庭環境で任意の物体を見つけて任意の場所に置くオープン語彙移動マニピュレーションのベンチマークを提案し、NeurIPS 2023コンペを実施。シミュレーションと実世界で評価し、成功チームに共通するエラー検出・回復と知覚・意思決定の統合の重要性を示した。
- 人間とエージェントの協調における高速オンライン適応のための線形モデルのブートストラップ人間エージェント協調2024/4/1
大規模非線形モデルで低容量のロジスティック回帰モデルを初期化し、協調中のオンライン更新を高速化する手法BLR-HACを提案。シミュレーションの物体再配置タスクで、少ない計算量で大規模モデル並みの性能を実現した。
- 基盤モデルの実世界ロボット応用:レビュー基盤モデル応用2024/2/1
LLMやVLMなどの基盤モデルを実ロボットシステムの知覚・運動計画・制御にどう組み込むかを、入出力関係の観点から整理したレビュー。
- OK-Robot: What Really Matters in Integrating Open-Knowledge Models for Robotics2024/1/1
- GOAT: GO to Any Thing2023/11/1
- HomeRobot: Open-Vocabulary Mobile Manipulation2023/6/1
- Evaluating Continual Learning on a Home Robot2023/6/1
- HACMan: Learning Hybrid Actor-Critic Maps for 6D Non-Prehensile Manipulation2023/5/1
- USA-Net: Unified Semantic and Affordance Representations for Robot Memory2023/4/1
- Navigating to Objects Specified by Images2023/4/1
- Spatial-Language Attention Policies for Efficient Robot Learning2023/4/1
- Task and Motion Planning with Large Language Models for Object Rearrangement2023/3/1
- StructDiffusion: Language-Guided Creation of Physically-Valid Structures using Unseen Objects2022/11/1
- CLIP-Fields: Weakly Supervised Semantic Fields for Robotic Memory2022/10/1
- HandoverSim: A Simulation Framework and Benchmark for Human-to-Robot Object Handovers2022/5/1
- Correcting Robot Plans with Natural Language Feedback2022/4/1
- Model Predictive Control for Fluid Human-to-Robot Handovers2022/4/1
- IFOR: Iterative Flow Minimization for Robotic Object Rearrangement2022/2/1
- Transporters with Visual Foresight for Solving Unseen Rearrangement Tasks2022/2/1
- Optimizing robot planning domains to reduce search time for long-horizon planning2021/11/1
- Learning Perceptual Concepts by Bootstrapping from Human Queries2021/11/1
- StructFormer: Learning Spatial Structure for Language-Guided Semantic Rearrangement of Novel Objects2021/10/1
- SORNet: Spatial Object-Centric Representations for Sequential Manipulation2021/9/1
- Predicting Stable Configurations for Semantic Placement of Novel Objects2021/8/1
- A Persistent Spatial Semantic Representation for High-level Natural Language Instruction Execution2021/7/1
- Language Grounding with 3D Objects2021/7/1
- NeRP: Neural Rearrangement Planning for Unknown Objects2021/6/1
- Automated Generation of Robotic Planning Domains from Observations2021/5/1
- Alternative Paths Planner (APP) for Provably Fixed-time Manipulation Planning in Semi-structured Environments2020/12/1
- Reactive Human-to-Robot Handovers of Arbitrary Objects2020/11/1
- Reactive Long Horizon Task Execution via Visual Skill and Precondition Models2020/11/1
- Transferable Task Execution from Pixels through Deep Planning Domain Learning2020/3/1
- Human Grasp Classification for Reactive Human-to-Robot Handovers2020/3/1
- 6-DOF Grasping for Target-driven Object Manipulation in Clutter2019/12/1
- Online Replanning in Belief Space for Partially Observable Task and Motion Problems2019/11/1
- Motion Reasoning for Goal-Based Imitation Learning2019/11/1
- Conditional Driving from Natural Language Instructions2019/10/1
- Collaborative Behavior Models for Optimized Human-Robot Teamwork2019/10/1
- "Good Robot!": Efficient Reinforcement Learning for Multi-Step Visual Tasks with Sim to Real Transfer2019/9/1
- Representing Robot Task Plans as Robust Logical-Dynamical Systems2019/8/1
- Prospection: Interpretable Plans From Language By Predicting the Future2019/3/1
- Evaluating Methods for End-User Creation of Robot Task Plans2018/11/1
- The CoSTAR Block Stacking Dataset: Learning with Workspace Constraints2018/10/1
- Visual Robot Task Planning2018/4/1
- Occupancy Map Prediction Using Generative and Fully Convolutional Networks for Vehicle Navigation2018/3/1
- Learning to Imagine Manipulation Goals for Robot Task Planning2017/11/1
- Temporal and Physical Reasoning for Perception-Based Robotic Manipulation2017/10/1
- User Experience of the CoSTAR System for Instruction of Collaborative Robots2017/3/1
- Combining Neural Networks and Tree Search for Task and Motion Planning in Challenging Environments2017/3/1
- Do What I Want, Not What I Did: Imitation of Skills by Planning Sequences of Actions2016/12/1
- CoSTAR: Instructing Collaborative Robots with Behavior Trees and Vision2016/11/1
- Towards Robot Task Planning From Probabilistic Models of Human Skills2016/2/1