Pengxiang Ding
収録論文 39本 ・ フィジカルAI/ロボット学習
評価基盤ロボットポリシー評価
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- XPolicyLab: ロボットポリシー評価と展開のための統一標準・オープンエコシステム評価基盤2026/8/10
ロボットポリシーの評価と展開を統一する標準規格とオープンエコシステムを提案し、N個のポリシーとM個の環境の接続コストをO(NM)からO(N+M)に削減する。
- XPolicyLab: ロボットポリシー評価と展開のための統一標準・オープンエコシステムロボットポリシー評価2026/8/10
ロボットポリシーの評価と展開を統一する標準規格とオープンエコシステムを提案し、N個のポリシーとM個の環境の接続コストをO(NM)からO(N+M)に削減した。
- XPolicyLab: A Unified Standard and Open Ecosystem for Robot Policy Evaluation and Deployment2026/8/1
- Revisiting Embodied Chain-of-Thought for Generalizable Robot Manipulation2026/6/1
- CapVector: Learning Transferable Capability Vectors in Parametric Space for Vision-Language-Action Models2026/5/1
- RoboMemArena: A Comprehensive and Challenging Robotic Memory Benchmark2026/5/1
- CUBic: Coordinated Unified Bimanual Perception and Control Framework2026/5/1
- Fast-dVLA: Accelerating Discrete Diffusion VLA to Real-Time Performance2026/3/1
- VAMPO: Policy Optimization for Improving Visual Dynamics in Video Action Models2026/3/1
- MMaDA-VLA: Large Diffusion Vision-Language-Action Model with Unified Multi-Modal Instruction and Generation2026/3/1
- Rethinking the Practicality of Vision-language-action Model: A Comprehensive Benchmark and An Improved Baseline2026/2/1
- HiF-VLA: Hindsight, Insight and Foresight through Motion Representation for Vision-Language-Action Models2025/12/1
- Embodied Robot Manipulation in the Era of Foundation Models: Planning and Learning Perspectives2025/12/1
- Unified Diffusion VLA: Vision-Language-Action Model via Joint Discrete Denoising Diffusion Process2025/11/1
- VCoT-Grasp: Grasp Foundation Models with Visual Chain-of-Thought Reasoning for Language-driven Grasp Generation2025/10/1
- VLA-RFT: Vision-Language-Action Reinforcement Fine-tuning with Verified Rewards in World Simulators2025/10/1
- Spatial Forcing: Implicit Spatial Representation Alignment for Vision-language-action Model2025/10/1
- VLA^2: Empowering Vision-Language-Action Models with an Agentic Framework for Unseen Concept Manipulation2025/10/1
- Towards a Unified Understanding of Robot Manipulation: A Comprehensive Survey2025/10/1
- VLA-Adapter: An Effective Paradigm for Tiny-Scale Vision-Language-Action Model2025/9/1
- Robust Online Residual Refinement via Koopman-Guided Dynamics Modeling2025/9/1
- TrajBooster: Boosting Humanoid Whole-Body Manipulation via Trajectory-Centric Learning2025/9/1
- Long-VLA: Unleashing Long-Horizon Capability of Vision Language Action Model for Robot Manipulation2025/8/1
- ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver2025/8/1
- CEED-VLA: Consistency Vision-Language-Action Model with Early-Exit Decoding2025/6/1
- ReinboT: Amplifying Robot Visual-Language Manipulation with Reinforcement Learning2025/5/1
- Unveiling the Potential of Vision-Language-Action Models with Open-Ended Multimodal Instructions2025/5/1
- OpenHelix: A Short Survey, Empirical Analysis, and Open-Source Dual-System VLA Model for Robotic Manipulation2025/5/1
- MoRE: Unlocking Scalability in Reinforcement Learning for Quadruped Vision-Language-Action Models2025/3/1
- PD-VLA: Accelerating Vision-Language-Action Model Integrated with Action Chunking via Parallel Decoding2025/3/1
- Rethinking Latent Redundancy in Behavior Cloning: An Information Bottleneck Approach for Robot Manipulation2025/2/1
- VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation2025/2/1
- Humanoid-VLA: Towards Universal Humanoid Control with Visual Integration2025/2/1
- GEVRM: Goal-Expressive Video Generation Model For Robust Visual Manipulation2025/2/1
- CARP: Visuomotor Policy Learning via Coarse-to-Fine Autoregressive Prediction2024/12/1
- QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning2024/12/1
- Score and Distribution Matching Policy: Advanced Accelerated Visuomotor Policies via Matched Distillation2024/12/1
- GeRM: A Generalist Robotic Model with Mixture-of-experts for Quadruped Robot2024/3/1
- QUAR-VLA: Vision-Language-Action Model for Quadruped Robots2023/12/1