Yichen Zhu
収録論文 27本 ・ フィジカルAI/ロボット学習
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- ActiveMimic: Egocentric Video Pretraining with Active Perception2026/6/1
- Hi-WM: Human-in-the-World-Model for Scalable Robot Post-Training2026/4/1
- dWorldEval: Scalable Robotic Policy Evaluation via Discrete Diffusion World Model2026/4/1
- Drive-P2D: A Progressive Perception-to-Decision Benchmark for VLMs in Autonomous Driving2026/1/1
- Let Me Show You: Learning by Retrieving from Egocentric Video for Robotic Manipulation2025/11/1
- HumanoidExo: Scalable Whole-Body Humanoid Manipulation via Wearable Exoskeleton2025/10/1
- ActiveUMI: Robotic Manipulation with Active Perception from Robot-Free Human Demonstrations2025/10/1
- dVLA: Diffusion Vision-Language-Action Model with Multimodal Chain-of-Thought2025/9/1
- ChatVLA-2: Vision-Language-Action Model with Open-World Embodied Reasoning from Pretrained Knowledge2025/5/1
- WorldEval: World Model as Real-World Robot Policies Evaluator2025/5/1
- PointVLA: Injecting the 3D World into Vision-Language-Action Models2025/3/10
- PointVLA: Injecting the 3D World into Vision-Language-Action Models2025/3/1
- ChatVLA: Unified Multimodal Understanding and Robot Control with Vision-Language-Action Model2025/2/20
- DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control2025/2/1
- ObjectVLA: End-to-End Open-World Object Manipulation Without Demonstration2025/2/1
- ChatVLA: Unified Multimodal Understanding and Robot Control with Vision-Language-Action Model2025/2/1
- Diffusion-VLA: Generalizable and Interpretable Robot Foundation Model via Self-Generated Reasoning2024/12/1
- CoA-VLA: Improving Vision-Language-Action Models via Visual-Textual Chain-of-Affordance2024/12/1
- TinyVLA: Towards Fast, Data-Efficient Vision-Language-Action Models for Robotic Manipulation2024/9/1
- Discrete Policy: Learning Disentangled Action Space for Multi-Task Robotic Manipulation2024/9/1
- Scaling Diffusion Policy in Transformer to 1 Billion Parameters for Robotic Manipulation2024/9/1
- MMRo: Are Multimodal LLMs Eligible as the Brain for In-Home Robotics?2024/6/1
- Retrieval-Augmented Embodied Agents2024/4/1
- Object-Centric Instruction Augmentation for Robotic Manipulation2024/1/1
- Language-Conditioned Robotic Manipulation with Fast and Slow Thinking2024/1/1
- Visual Robotic Manipulation with Depth-Aware Pretraining2024/1/1
- Revisiting Event-based Video Frame Interpolation2023/7/1