Weinan Zhang
収録論文 45本 ・ フィジカルAI/ロボット学習
強化学習/ヒューマノイドシミュレーション評価
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- RoboStriker: 自律型ヒューマノイドボクシングのための潜在空間戦略ゲーム強化学習/ヒューマノイド2026/8/17
ヒューマノイドボクシングを2プレイヤーの潜在空間ゼロサムマルコフゲームとして定式化し、戦略的探索と物理的実現性の矛盾を解決する階層的フレームワークを提案した。
- GAUGE: 物理的忠実性を測定するための実世界基盤ベンチマークシミュレーション評価2026/8/6
シミュレーションエンジンと生成ビデオワールドモデルの物理的忠実性を、実世界の軌跡に基づいて診断するベンチマークを提案した。
- GAUGE: A Measurement-Grounded Benchmark for Physical Fidelity in Simulation Engines and Video World Models2026/8/1
- InternVLA-A1.5: Unifying Understanding, Latent Foresight, and Action for Compositional Generalization2026/7/1
- Scaling Behavior Foundation Model for Humanoid Robots2026/7/1
- HELP: Human-Efficient Large-Scale Robot Post-Training with Rollout Segmentation2026/7/1
- RoboSnap: One-Shot Real-to-Sim Scene Generation for Generalizable Robot Learning and Evaluation2026/7/1
- AHA-WAM:Asynchronous Horizon-Adaptive World-Action Modeling with Observation-Guided Context Routing2026/6/1
- EBench: Elemental Diagnosis of Generalist Mobile Manipulation Policies2026/6/1
- PRTS: A Primitive Reasoning and Tasking System via Contrastive Representations2026/4/1
- PEPA: a Persistently Autonomous Embodied Agent with Personalities2026/3/1
- Embodiment-Aware Generalist Specialist Distillation for Unified Humanoid Whole-Body Control2026/2/1
- Scalable and General Whole-Body Control for Cross-Humanoid Locomotion2026/2/1
- TextOp: Real-time Interactive Text-Driven Humanoid Robot Motion Generation and Control2026/2/1
- Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models2026/2/1
- UniCon: A Unified System for Efficient Robot Learning Transfers2026/1/1
- RoboStriker: Hierarchical Decision-Making for Autonomous Humanoid Boxing2026/1/1
- H-Zero: Cross-Humanoid Locomotion Pretraining Enables Few-shot Novel Embodiment Transfer2025/12/1
- RLinf-VLA: A Unified and Efficient Framework for Reinforcement Learning of Vision-Language-Action Models2025/10/1
- Manipulation as in Simulation: Enabling Accurate Geometry Perception in Robots2025/9/1
- TrajBooster: Boosting Humanoid Whole-Body Manipulation via Trajectory-Centric Learning2025/9/1
- KungfuBot2: Learning Versatile Motion Skills for Humanoid Whole-Body Control2025/9/1
- CookBench: A Long-Horizon Embodied Planning Benchmark for Complex Cooking Scenarios2025/8/1
- MobileUse: A GUI Agent with Hierarchical Reflection for Autonomous Mobile Operation2025/7/1
- UniTracker: Learning Universal Whole-Body Motion Tracker for Humanoid Robots2025/7/1
- KungfuBot: Physics-Based Humanoid Whole-Body Control for Learning Highly-Dynamic Skills2025/6/1
- MARFT: Multi-Agent Reinforcement Fine-Tuning2025/4/1
- PALo: Learning Posture-Aware Locomotion for Quadruped Robots2025/3/1
- DriveGen: Towards Infinite Diverse Traffic Scenarios with Large Models2025/3/1
- BeamDojo: Learning Agile Humanoid Locomotion on Sparse Footholds2025/2/1
- A Unified and General Humanoid Whole-Body Controller for Versatile Locomotion2025/2/1
- Humanoid Whole-Body Locomotion on Narrow Terrain via Dynamic Balance and Reinforcement Learning2025/2/1
- RHINO: Learning Real-Time Humanoid-Human-Object Interaction from Human Demonstrations2025/2/1
- Re$^3$Sim: Generating High-Fidelity Simulation Data via 3D-Photorealistic Real-to-Sim for Robotic Manipulation2025/2/1
- GenSim2: Scaling Robot Data Generation with Multi-modal and Reasoning LLMs2024/10/1
- LoopSR: Looping Sim-and-Real for Lifelong Policy Adaptation of Legged Robots2024/9/1
- World Model-based Perception for Visual Legged Locomotion2024/9/1
- OmniH2O: Universal and Dexterous Human-to-Humanoid Whole-Body Teleoperation and Learning2024/6/1
- Learning an Actionable Discrete Diffusion Policy via Large-Scale Actionless Video Pre-Training2024/2/1
- Vision-Language Foundation Models as Effective Robot Imitators2023/11/1
- Bridging the Sim-to-Real Gap from the Information Bottleneck Perspective2023/5/1
- Sim-to-Real Transfer for Quadrupedal Locomotion via Terrain Transformer2022/12/1
- Multi-embodiment Legged Robot Control as a Sequence Modeling Problem2022/12/1
- NeurIPS 2022 Competition: Driving SMARTS2022/11/1
- Bootstrapped Transformer for Offline Reinforcement Learning2022/6/1