Abhinav Gupta
Carnegie Mellon University
収録論文 72本 ・ フィジカルAI/ロボット学習
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- BRIDGE: 形態と制御の共設計による物理AIのためのオープンソース人型ロボットプラットフォーム人型ロボット/共設計2026/9/3
人間の動作データを活用できる人型ロボットの設計を、形態と全身制御をデータ駆動で共最適化するフレームワークを提案し、その設計を実現したオープンソースの小型人型ロボット「Bridge」を公開した。
- マイクロロボット向け高トルク密度PCBアキシャルフラックス永久磁石モータアクチュエータ2025/9/1
IC基板HDI技術で作製した48層PCB巻線により、直径19mm・厚さ5mmで銅充填率45%を達成したマイクロアキシャルフラックスモータを開発し、電磁・熱解析と試作機で性能を検証した。
- 進化的方策最適化強化学習2025/3/1
進化的アルゴリズムのスケーラビリティと方策勾配の安定性を組み合わせたハイブリッド強化学習手法EPOを提案し、巧みな操作や歩行タスクで性能を向上させた。
- Gen2Act: 人間動画生成を活用した汎用ロボットマニピュレーションマニピュレーション2024/9/1
ウェブデータで学習した人間動画生成モデルを使い、生成された動画を条件にロボット方策を学習することで、未見物体や新規動作への汎化を実現した。
- 長期的なソフトロボティクスデータ収集のためのモジュラーパラレルマニピュレータソフトロボティクス2024/9/1
ソフトロボットの大規模データ収集を可能にする、市販モータと柔軟なパラレル構造の指を組み合わせたモジュラーロボットプラットフォームを提案し、方策勾配強化学習への適用可能性を検証した。
- HRP: 人間のアフォーダンスを活用したロボット事前学習VLA2024/7/1
インターネット上の人間動画から手・物体・接触のアフォーダンスを自動抽出し、既存の視覚表現を微調整することで、多様なロボットタスクの性能を向上させる手法を提案。
- 触覚を聴く:接触の多いマニピュレーションのための音声視覚事前学習マニピュレーション2024/5/1
接触マイクを触覚センサとして用い、大規模な音声視覚データで事前学習することで、ロボットマニピュレーションの性能を向上させる手法を提案。
- Track2Act: インターネット動画からの点トラック予測による汎化可能なロボットマニピュレーションマニピュレーション2024/5/1
ウェブ上の人間やロボットの動画から画像内の点の動きを予測し、それを物体の剛体変換に変換してロボットの手先軌道を生成することで、未見の物体や場面でも適応なしに操作できる汎用的なマニピュレーション手法を提案した。
- DROID: 大規模実環境ロボットマニピュレーションデータセットマニピュレーション2024/3/1
北米・アジア・欧州の50名が12ヶ月かけて564シーン・84タスクで収集した76k軌道・350時間のロボット操作データセットを構築し、学習ポリシーの性能と汎化性能が向上することを示した。
- 連続的な系列間モデリングのための階層的状態空間モデル時系列モデリング2024/2/1
生のセンサ系列から物理量の系列を予測するため、構造化状態空間モデルを積み重ねた時間的階層モデルHiSSを提案し、6つの実世界データセットでTransformerやMamba等を上回る精度を示した。
- Towards Generalizable Zero-Shot Manipulation via Translating Human Interaction Plans2023/12/1
- Exploitation-Guided Exploration for Semantic Embodied Navigation2023/11/1
- Open X-Embodiment: Robotic Learning Datasets and RT-X Models2023/10/1
- An Unbiased Look at Datasets for Visuo-Motor Pre-Training2023/10/1
- RoboAgent: Generalization and Efficiency in Robot Manipulation via Semantic Augmentations and Action Chunking2023/9/1
- Evaluating Continual Learning on a Home Robot2023/6/1
- Train Offline, Test Online: A Real Robot Learning Benchmark2023/6/1
- Visual Affordance Prediction for Guiding Robot Exploration2023/5/1
- Affordance Diffusion: Synthesizing Hand-Object Interactions2023/3/1
- Manipulate by Seeing: Creating Manipulation Controllers from Pre-Trained Representations2023/3/1
- Zero-Shot Robot Manipulation from Passive Human Videos2023/2/1
- Design of an All-Purpose Terrace Farming Robot2022/12/1
- Last-Mile Embodied Visual Navigation2022/11/1
- Real World Offline Reinforcement Learning with Realistic Data Source2022/10/1
- All the Feels: A dexterous hand with large-area tactile sensing2022/10/1
- Learning Dexterous Manipulation from Exemplar Object Trajectories and Pre-Grasps2022/9/1
- Human-to-Robot Imitation in the Wild2022/7/1
- Can Foundation Models Perform Zero-Shot Task Specification For Robot Manipulation?2022/4/1
- R3M: A Universal Visual Representation for Robot Manipulation2022/3/1
- RB2: Robotic Manipulation Benchmarking with a Twist2022/3/1
- The Unsurprising Effectiveness of Pre-Trained Vision Models for Control2022/3/1
- A Differentiable Recipe for Learning Visual Non-Prehensile Planar Manipulation2021/11/1
- ReSkin: versatile, replaceable, lasting tactile skins2021/11/1
- Learning Multi-Objective Curricula for Robotic Policy Learning2021/10/1
- The Functional Correspondence Problem2021/9/1
- Hierarchical Neural Dynamic Policies2021/7/1
- DeepMPCVS: Deep Model Predictive Control for Visual Servoing2021/5/1
- Learn-to-Race: A Multimodal Control Environment for Autonomous Racing2021/3/1
- droidlet: modular, heterogenous, multi-modal agents2021/1/1
- Where2Act: From Pixels to Actions for Articulated 3D Objects2021/1/1
- Neural Dynamic Policies for End-to-End Sensorimotor Learning2020/12/1
- Same Object, Different Grasps: Data and Semantic Knowledge for Task-Oriented Grasping2020/11/1
- Transformers for One-Shot Visual Imitation2020/11/1
- Visual Imitation Made Easy2020/8/1
- Swoosh! Rattle! Thump! -- Actions that Sound2020/7/1
- Object Goal Navigation using Goal-Oriented Semantic Exploration2020/7/1
- Learning Robot Skills with Temporal Variational Inference2020/6/1
- Neural Topological SLAM for Visual Navigation2020/5/1
- Learning to Explore using Active Neural SLAM2020/4/1
- Use the Force, Luke! Learning to Predict Physical Forces by Simulating Effects2020/3/1
- Third-Person Visual Imitation Learning via Decoupled Hierarchical Controller2019/11/1
- Object-centric Forward Modeling for Model Predictive Control2019/10/1
- Efficient Bimanual Manipulation Using Learned Task Schemas2019/9/1
- Environment Probing Interaction Policies2019/7/1
- PyRobot: An Open-source Robotics Framework for Research and Benchmarking2019/6/1
- Self-Supervised Exploration via Disagreement2019/6/1
- Learning Exploration Policies for Navigation2019/3/1
- Hardware Conditioned Policies for Multi-Robot Transfer Learning2018/11/1
- Visual Semantic Navigation using Scene Priors2018/10/1
- Multiple Interactions Made Easy (MIME): Large Scale Demonstrations Data for Imitation2018/10/1
- Robot Learning in Homes: Improving Generalization and Reducing Dataset Bias2018/7/1
- Learning to Grasp Without Seeing2018/5/1
- Learning 6-DOF Grasping Interaction via Deep Geometry-aware 3D Representations2017/8/1
- CASSL: Curriculum Accelerated Self-Supervised Learning2017/8/1
- Visual Semantic Planning using Deep Successor Representations2017/5/1
- Learning to Fly by Crashing2017/4/1
- Robust Adversarial Reinforcement Learning2017/3/1
- PixelNet: Representation of the pixels, by the pixels, and for the pixels2017/2/1
- Supervision via Competition: Robot Adversaries for Learning Tasks2016/10/1
- Learning to Push by Grasping: Using multiple tasks for effective learning2016/9/1
- The Curious Robot: Learning Visual Representations via Physical Interactions2016/4/1
- Supersizing Self-supervision: Learning to Grasp from 50K Tries and 700 Robot Hours2015/9/1