Zeyu Zhang
Shanghai Jiao Tong University
収録論文 32本 ・ フィジカルAI/ロボット学習
ビデオ生成/ベンチマーク
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- H2R-Bench: 世界モデルにおける人間からロボットへの操作ビデオ生成のベンチマークビデオ生成/ベンチマーク2026/8/13
人間の操作ビデオをロボットの操作ビデオに変換する能力を評価するベンチマークを提案し、既存のビデオ生成モデルの限界を明らかにした。
- H2R-Bench: Benchmarking Human-to-Robot Manipulation Video Generation in World Models2026/8/1
- DragMesh-2: Physically Plausible Dexterous Hand-Object Interaction with Articulated Objects2026/6/1
- MotionVLA: Vision-Language-Action Model for Humanoid Motion2026/6/1
- GeneralVLA-2: Geometry-Aware Reconstruction and Governed Memory for Robot Planning2026/6/1
- MWM: Mobile World Models for Action-Conditioned Consistent Prediction2026/3/1
- GeneralVLA: Generalizable Vision-Language-Action Models with Knowledge-Guided Trajectory Planning2026/2/1
- GeoWorld: Geometric World Models2026/2/1
- MobileVLA-R1: Reinforcing Vision-Language-Action for Mobile Robots2025/11/1
- VLA-R1: Enhancing Reasoning in Vision-Language-Action Models2025/10/1
- StereoAdapter: Adapting Stereo Depth Estimation to Underwater Scenes2025/9/1
- Nav-R1: Reasoning and Navigation in Embodied Scenes2025/9/1
- Integration of Robot and Scene Kinematics for Sequential Mobile Manipulation Planning2025/8/1
- SuperMag: Vision-based Tactile Data Guided High-resolution Tactile Shape Reconstruction for Magnetic Tactile Sensors2025/7/1
- IKDiffuser: a Diffusion-based Generative Inverse Kinematics Solver for Kinematic Trees2025/6/1
- ManipLVM-R1: Reinforcement Learning for Reasoning in Embodied Manipulation with Large Vision-Language Models2025/5/1
- A2I-Calib: An Anti-noise Active Multi-IMU Spatial-temporal Calibration Framework for Legged Robots2025/3/1
- Hazards in Daily Life? Enabling Robots to Proactively Detect and Resolve Anomalies2024/11/1
- M2Diffuser: Diffusion-based Trajectory Optimization for Mobile Manipulation in 3D Scenes2024/10/1
- M3Bench: Benchmarking Whole-body Motion Generation for Mobile Manipulation in 3D Scenes2024/10/1
- PR2: A Physics- and Photo-realistic Humanoid Testbed with Pilot Study in Competition2024/9/1
- Flight Structure Optimization of Modular Reconfigurable UAVs2024/7/1
- Closed-Loop Open-Vocabulary Mobile Manipulation with GPT-4V2024/4/1
- LLM3:Large Language Model-based Task and Motion Planning with Motion Failure Reasoning2024/3/1
- Part-level Scene Reconstruction Affords Robot Interaction2023/7/1
- A Reconfigurable Data Glove for Reconstructing Physical and Virtual Grasps2023/1/1
- Sequential Manipulation Planning on Scene Graph2022/7/1
- Understanding Physical Effects for Effective Tool-use2022/6/1
- Efficient Task Planning for Mobile Manipulation: a Virtual Kinematic Chain Perspective2021/8/1
- Consolidating Kinematic Models to Promote Coordinated Mobile Manipulations2021/8/1
- Reconstructing Interactive 3D Scenes by Panoptic Mapping and CAD Model Alignments2021/3/1
- Human-Robot Interaction in a Shared Augmented Reality Workspace2020/7/1