Dhruv Shah
収録論文 60本 ・ フィジカルAI/ロボット学習
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- Emergent Compositional Skills in Mixture-of-Experts VLAs2026/7/1
- RoboVista: Evaluating Vision Language Models for Diverse Robot Applications2026/7/1
- Addressing the Orchestration Gap in Generalist Robots via Physical Agency2026/7/1
- What Matters in Orchestrating Robot Policies: A Systematic Study of Hierarchical VLA Agents2026/6/9
- What Matters in Orchestrating Robot Policies: A Systematic Study of Hierarchical VLA Agents2026/6/1
- RAVEN: Long-Horizon Reasoning & Navigation with a Visuo-Spatio-Temporal Memory2026/6/1
- Visual Verification Enables Inference-time Steering and Autonomous Policy Improvement2026/6/1
- WorldArena 2.0: Extending Embodied World Model Benchmarking on Modality, Functionality and Platform2026/5/1
- Grounding Robot Generalization in Training Data via Retrieval-Augmented VLMs2026/3/1
- MolmoB0T: Large-Scale Simulation Enables Zero-Shot Manipulation2026/3/1
- WorldArena: A Unified Benchmark for Evaluating Perception and Functional Utility of Embodied World Models2026/2/1
- BPP: Long-Context Robot Imitation Learning by Focusing on Key History Frames2026/2/1
- LAP: Language-Action Pre-Training Enables Zero-shot Cross-Embodiment Transfer2026/2/1
- AsyncVLA: An Asynchronous VLA for Fast and Robust Navigation on the Edge2026/2/1
- PolaRiS: Scalable Real-to-Sim Evaluations for Generalist Robot Policies2025/12/1
- Evaluating Gemini Robotics Policies in a Veo World Simulator2025/12/1
- Text to Robotic Assembly of Multi Component Objects using 3D Generative AI and Vision Language Models2025/11/1
- What Matters in RL-Based Methods for Object-Goal Navigation? An Empirical Study and A Unified Framework2025/10/1
- Gemini Robotics 1.5: Pushing the Frontier of Generalist Robots with Advanced Embodied Reasoning, Thinking, and Motion Transfer2025/10/1
- OmniVLA: An Omni-Modal Vision-Language-Action Model for Robot Navigation2025/9/23
- OmniVLA: An Omni-Modal Vision-Language-Action Model for Robot Navigation2025/9/1
- Towards Data-Driven Metrics for Social Robot Navigation Benchmarking2025/9/1
- CAST: Counterfactual Labels Improve Instruction Following in Vision-Language-Action Models2025/8/1
- Bridging Perception and Action: Spatially-Grounded Mid-Level Representations for Robot Generalization2025/6/1
- Guiding Data Collection via Factored Scaling Curves2025/5/1
- Learning to Drive Anywhere with Model-Based Reannotation2025/5/1
- A Taxonomy for Evaluating Generalist Robot Manipulation Policies2025/3/3
- Gemini Robotics: Bringing AI into the Physical World2025/3/1
- A Taxonomy for Evaluating Generalist Robot Manipulation Policies2025/3/1
- Robot Data Curation with Mutual Information Estimators2025/2/1
- STEER: Flexible Robotic Manipulation via Dense Language Grounding2024/11/1
- Vision Language Models are In-Context Value Learners2024/11/1
- LeLaN: Learning A Language-Conditioned Navigation Policy from In-the-Wild Videos2024/10/1
- Traversability-Aware Legged Navigation by Learning from Real-World Visual Data2024/10/1
- Gen2Act: Human Video Generation in Novel Scenarios enables Generalizable Robot Manipulation2024/9/1
- Mobility VLA: Multimodal Instruction Navigation with Long-Context VLMs and Topological Graphs2024/7/1
- SELFI: Autonomous Self-Improvement with Reinforcement Learning for Social Navigation2024/3/1
- Pushing the Limits of Cross-Embodiment Learning for Manipulation and Navigation2024/2/1
- Bridging Language and Action: A Survey of Language-Conditioned Robot Manipulation2023/12/1
- GOAT: GO to Any Thing2023/11/1
- Open X-Embodiment: Robotic Learning Datasets and RT-X Models2023/10/1
- NoMaD: Goal Masked Diffusion Policies for Navigation and Exploration2023/10/1
- Navigation with Large Language Models: Semantic Guesswork as a Heuristic for Planning2023/10/1
- ViNT: A Foundation Model for Visual Navigation2023/6/1
- SACSoN: Scalable Autonomous Control for Social Navigation2023/6/1
- FastRLAP: A System for Learning High-Speed Driving via Deep RL and Autonomous Practicing2023/4/1
- Grounded Decoding: Guiding Text Generation with Grounded Models for Embodied Agents2023/3/1
- Offline Reinforcement Learning for Visual Navigation2022/12/1
- Learning Robotic Navigation from Experience: Principles, Methods, and Recent Results2022/12/1
- ExAug: Robot-Conditioned Navigation Policies via Geometric Experience Augmentation2022/10/1
- GNM: A General Navigation Model to Drive Any Robot2022/10/1
- LM-Nav: Robotic Navigation with Large Pre-Trained Models of Language, Vision, and Action2022/7/1
- ViKiNG: Vision-Based Kilometer-Scale Navigation with Geographic Hints2022/2/1
- Value Function Spaces: Skill-Centric State Abstractions for Long-Horizon Reasoning2021/11/1
- Hybrid Imitative Planning with Geometric and Predictive Costs in Off-road Environments2021/11/1
- Rapid Exploration for Open-World Navigation with Latent Goal Models2021/4/1
- ViNG: Learning Open-World Navigation with Visual Goals2020/12/1
- Aerial Manipulation Using Hybrid Force and Position NMPC Applied to Aerial Writing2020/6/1
- The Ingredients of Real-World Robotic Reinforcement Learning2020/4/1
- Robust Localization of an Arbitrary Distribution of Radioactive Sources for Aerial Inspection2017/10/1