Yan Ding
収録論文 40本 ・ フィジカルAI/ロボット学習
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- InternVLA-A1.5: Unifying Understanding, Latent Foresight, and Action for Compositional Generalization2026/7/1
- One-to-Two Acting: A Novel Framework for Single-arm Agent Action Expansion to Dual Arms2026/6/1
- A Scalable Embodied Intelligence Platform for Seamless Real-to-Sim-to-Real Transfer of Household Mobile Manipulation Tasks2026/6/1
- UMI-Bench 1.0: An Open and Reproducible Real-World Benchmark for Tabletop Robotic Manipulation with UMI Data2026/6/1
- VISTA: Vision-Grounded and Physics-Validated Adaptation of UMI data for VLA Training2026/6/1
- EgoAERO: Learning Dexterous Manipulation from a Single Egocentric Video without Object Assets2026/6/1
- GSAM: A Generalizable and Safe Robotic Framework for Articulated Object Manipulation2026/5/1
- Robot Planning and Situation Handling with Active Perception2026/4/1
- AsyncVLA: Asynchronous Flow Matching for Vision-Language-Action Models2025/11/18
- LLM-GROP: Visually Grounded Robot Task and Motion Planning with Large Language Models2025/11/1
- AsyncVLA: Asynchronous Flow Matching for Vision-Language-Action Models2025/11/1
- FastUMI-100K: Advancing Data-driven Robotic Manipulation with a Large-scale UMI-style Dataset2025/10/1
- U-ARM : Ultra low-cost general teleoperation interface for robot manipulation2025/9/1
- MLM: Learning Multi-task Loco-Manipulation Whole-Body Control for Quadruped Robot with Arm2025/8/1
- Skill-Nav: Enhanced Navigation with Versatile Quadrupedal Locomotion via Waypoint Interface2025/6/1
- Cross from Left to Right Brain: Adaptive Text Dreamer for Vision-and-Language Navigation2025/5/1
- Think Small, Act Big: Primitive Prompt Learning for Lifelong Robot Manipulation2025/4/1
- MoMa-Kitchen: A 100K+ Benchmark for Affordance-Grounded Last-Mile Navigation in Mobile Manipulation2025/3/1
- AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems2025/3/1
- Openfly: A comprehensive platform for aerial vision-language navigation2025/2/1
- SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model2025/1/1
- BestMan: A Modular Mobile Manipulator Platform for Embodied AI with Unified Simulation-Hardware APIs2024/10/1
- AlignBot: Aligning VLM-powered Customized Task Planning with User Reminders Through Fine-Tuning for Household Robots2024/9/1
- FastUMI: A Scalable and Hardware-Independent Universal Manipulation Interface with Dataset2024/9/1
- A New Clustering-based View Planning Method for Building Inspection with Drone2024/8/1
- DKPROMPT: Domain Knowledge Prompting Vision-Language Models for Open-World Planning2024/6/1
- A Survey of Optimization-based Task and Motion Planning: From Classical To Learning Approaches2024/4/1
- MoMa-Pos: An Efficient Object-Kinematic-Aware Base Placement Optimization Framework for Mobile Manipulation2024/3/1
- ORLA*: Mobile Manipulator-Based Object Rearrangement with Lazy A Star2023/9/1
- Symbolic State Space Optimization for Long Horizon Mobile Manipulation Planning2023/7/1
- Integrating Action Knowledge and LLMs for Task Planning and Situation Handling in Open Worlds2023/5/1
- ARDIE: AR, Dialogue, and Eye Gaze Policies for Human-Robot Collaboration2023/5/1
- Grounding Classical Task Planners via Vision-Language Models2023/4/1
- Task and Motion Planning with Large Language Models for Object Rearrangement2023/3/1
- Robot Task Planning and Situation Handling in Open Worlds2022/10/1
- GLAD: Grounded Layered Autonomous Driving for Complex Service Tasks2022/10/1
- Learning to Ground Objects for Robot Task and Motion Planning2022/2/1
- Visually Grounded Task and Motion Planning for Mobile Manipulation2022/2/1
- Task-Motion Planning for Safe and Efficient Urban Driving2020/3/1
- Visual Semantic SLAM with Landmarks for Large-Scale Outdoor Environment2020/1/1