Xuelong Li
Institute of Artificial Intelligence, China Telecom
収録論文 64本 ・ フィジカルAI/ロボット学習
VLA/ナビゲーション
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- WNM-3D: 3Dシーン条件付けによるクローズドループVLNのための世界ナビゲーションモデルVLA/ナビゲーション2026/8/7
連続的な視覚言語ナビゲーション(VLN)のための生成的世界行動モデルを提案し、3Dシーン表現を条件として将来の視覚と行動を同時生成することで、クローズドループのナビゲーション性能を向上させた。
- WNM-3D: A World Navigation Model with 3D Scene Conditioning for Closed-Loop VLN2026/8/1
- RRTrack: Robust and Recoverable Object 6D Pose Tracking for Dynamic Scenes2026/7/1
- KineBench: Benchmarking Embodied World Models via IDM-Free Kinematic Grounding2026/7/1
- EDAR: Learning Environment-Dependent Action Representations for Robotic Manipulation2026/7/1
- From World Action Models to Embodied Brains: A Roadmap for Open-World Physical Intelligence2026/7/1
- GN0: Toward a Unified Paradigm for Generation, Evaluation, and Policy Learning in Visual-Language Navigation2026/6/1
- OASIS: From Simulation Data Collection to Real-World Humanoid Loco-Manipulation2026/6/1
- SpaceVLN: A Zero-Shot Vision-and-Language Navigation Agent with Online Spatial Cognitive Memory and Reasoning2026/6/1
- VISTA: Vision-Grounded and Physics-Validated Adaptation of UMI data for VLA Training2026/6/1
- PRTS: A Primitive Reasoning and Tasking System via Contrastive Representations2026/4/30
- PRTS: A Primitive Reasoning and Tasking System via Contrastive Representations2026/4/1
- Learn Weightlessness: Imitate Non-Self-Stabilizing Motions on Humanoid Robot2026/4/1
- RPG: Robust Policy Gating for Smooth Multi-Skill Transitions in Humanoid Fighting2026/4/1
- DeCoNav: Dialog enhanced Long-Horizon Collaborative Vision-Language Navigation2026/4/1
- Closed-Loop Action Chunks with Dynamic Corrections for Training-Free Diffusion Policy2026/3/1
- Pro-HOI: Perceptive Root-guided Humanoid-Object Interaction2026/3/1
- ZeroWBC: Learning Natural Whole-Body Humanoid Interaction from Human Egocentric Data2026/3/1
- X-Loco: Towards Generalist Humanoid Locomotion Control via Synergetic Policy Distillation2026/3/1
- PCHC: Enabling Preference Conditioned Humanoid Control via Multi-Objective Reinforcement Learning2026/3/1
- Beyond Short-Horizon: VQ-Memory for Robust Long-Horizon Manipulation in Non-Markovian Simulation Benchmarks2026/3/1
- HUSKY: Humanoid Skateboarding System via Physics-Aware Whole-Body Control2026/2/1
- SAGE-LLM: Towards Safe and Generalizable LLM Controller with Fuzzy-CBF Verification and Graph-Structured Knowledge Retrieval for UAV Decision2026/2/1
- TextOp: Real-time Interactive Text-Driven Humanoid Robot Motion Generation and Control2026/2/1
- Learning Soccer Skills for Humanoid Robots: A Progressive Perception-Action Framework2026/2/1
- The RoboSense Challenge: Sense Anything, Navigate Anywhere, Adapt Across Platforms2026/1/1
- Steering Vision-Language-Action Models as Anti-Exploration: A Test-Time Scaling Approach2025/12/1
- FastUMI-100K: Advancing Data-driven Robotic Manipulation with a Large-scale UMI-style Dataset2025/10/1
- CompassNav: Steering From Path Imitation To Decision Understanding In Navigation2025/10/1
- Towards Reliable LLM-based Robot Planning via Combined Uncertainty Estimation2025/10/1
- Align-Then-stEer: Adapting the Vision-Language Action Models through Unified Latent Guidance2025/9/1
- MLM: Learning Multi-task Loco-Manipulation Whole-Body Control for Quadruped Robot with Arm2025/8/1
- EO-1: An Open Unified Embodied Foundation Model for General Robot Control2025/8/1
- MoRE: Mixture of Residual Experts for Humanoid Lifelike Gaits Learning on Complex Terrains2025/6/1
- KungfuBot: Physics-Based Humanoid Whole-Body Control for Learning Highly-Dynamic Skills2025/6/1
- Dynamic Mixture of Progressive Parameter-Efficient Expert Library for Lifelong Robot Learning2025/6/1
- Dynamic Manipulation of Deformable Objects in 3D: Simulation, Benchmark and Learning Strategy2025/5/1
- Hume: Introducing System-2 Thinking in Visual-Language-Action Model2025/5/1
- Cross from Left to Right Brain: Adaptive Text Dreamer for Vision-and-Language Navigation2025/5/1
- Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction2025/5/1
- AutoBio: A Simulation and Benchmark for Robotic Automation in Digital Biology Laboratory2025/5/1
- Think Small, Act Big: Primitive Prompt Learning for Lifelong Robot Manipulation2025/4/1
- Adversarial Locomotion and Motion Imitation for Humanoid Policy Learning2025/4/1
- MoMa-Kitchen: A 100K+ Benchmark for Affordance-Grounded Last-Mile Navigation in Mobile Manipulation2025/3/1
- Openfly: A comprehensive platform for aerial vision-language navigation2025/2/1
- Humanoid Whole-Body Locomotion on Narrow Terrain via Dynamic Balance and Reinforcement Learning2025/2/1
- Leader and Follower: Interactive Motion Generation under Trajectory Constraints2025/2/1
- SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model2025/1/1
- Autonomous Decision Making for UAV Cooperative Pursuit-Evasion Game with Reinforcement Learning2024/11/1
- G3Flow: Generative 3D Semantic Flow for Pose-aware and Generalizable Object Manipulation2024/11/1
- Preference Aligned Diffusion Planner for Quadrupedal Locomotion Control2024/10/1
- AlignBot: Aligning VLM-powered Customized Task Planning with User Reminders Through Fine-Tuning for Household Robots2024/9/1
- FastUMI: A Scalable and Hardware-Independent Universal Manipulation Interface with Dataset2024/9/1
- COHERENT: Collaboration of Heterogeneous Multi-Robot System with Large Language Models2024/9/1
- KOI: Accelerating Online Imitation Learning via Hybrid Key-state Guidance2024/8/1
- Play to the Score: Stage-Guided Dynamic Multi-Sensory Fusion for Robotic Manipulation2024/8/1
- Depth Helps: Improving Pre-trained RGB-based Policy with Depth Information Injection2024/8/1
- Learning Manipulation by Predicting Interaction2024/6/1
- Towards Efficient LLM Grounding for Embodied Multi-Agent Collaboration2024/5/1
- SAM-E: Leveraging Visual Foundation Model with Sequence Imitation for Embodied Manipulation2024/5/1
- Learning an Actionable Discrete Diffusion Policy via Large-Scale Actionless Video Pre-Training2024/2/1
- Kinematic-aware Prompting for Generalizable Articulated Object Manipulation with LLMs2023/11/1
- Affordance-Driven Next-Best-View Planning for Robotic Grasping2023/9/1
- Robust Quadrupedal Locomotion via Risk-Averse Policy Learning2023/8/1