Jiazhao Zhang
収録論文 32本 ・ フィジカルAI/ロボット学習
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- ReferTrack: Referring Then Tracking for Embodied Visual Tracking2026/7/1
- Qwen-RobotManip Technical Report: Alignment Unlocks Scale for Robotic Manipulation Foundation Models2026/6/16
- DynaMOMA: Instantaneous Prediction of Grasp Poses for Mobile Manipulation of Dynamic Objects2026/6/1
- PAIWorld: A 3D-Consistent World Foundation Model for Robotic Manipulation2026/6/1
- AllDayNav: Lifelong Navigation via Real-World Reinforcement Learning2026/6/1
- Qwen-RobotNav Technical Report: A Scalable Navigation Model Designed for an Agentic Navigation System2026/6/1
- Qwen-RobotManip Technical Report: Alignment Unlocks Scale for Robotic Manipulation Foundation Models2026/6/1
- Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments2026/5/1
- Habitat-GS: A High-Fidelity Navigation Simulator with Dynamic Gaussian Splatting2026/4/1
- SPAN-Nav: Generalized Spatial Awareness for Versatile Vision-Language Navigation2026/3/1
- NavGSim: High-Fidelity Gaussian Splatting Simulator for Large-Scale Navigation2026/3/1
- LDA-1B: Scaling Latent Dynamics Action Model via Universal Embodied Data Ingestion2026/2/1
- ReWorld: Multi-Dimensional Reward Modeling for Embodied World Models2026/1/1
- AdaPower: Specializing World Foundation Models for Predictive Manipulation2025/12/1
- UrbanVLA: A Vision-Language-Action Model for Urban Micromobility2025/10/1
- MM-Nav: Multi-View VLA Model for Robust Visual Navigation via Multi-Expert Learning2025/10/1
- TrackVLA++: Unleashing Reasoning and Memory Capabilities in VLA Models for Embodied Visual Tracking2025/10/1
- Embodied Navigation Foundation Model2025/9/1
- DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge2025/7/1
- OctoNav: Towards Generalist Embodied Navigation2025/6/1
- LaDi-WM: A Latent Diffusion-based World Model for Predictive Manipulation2025/5/1
- TrackVLA: Embodied Visual Tracking in the Wild2025/5/1
- RoboVerse: Towards a Unified Platform, Dataset and Benchmark for Scalable and Generalizable Robot Learning2025/4/1
- SoFar: Language-Grounded Orientation Bridges Spatial Reasoning and Object Manipulation2025/2/1
- CogNav: Cognitive Process Modeling for Object Goal Navigation with LLMs2024/12/1
- Uni-NaVid: A Video-based Vision-Language-Action Model for Unifying Embodied Navigation Tasks2024/12/1
- GAPartManip: A Large-scale Part-centric Dataset for Material-Agnostic Articulated Object Manipulation2024/11/1
- InstruGen: Automatic Instruction Generation for Vision-and-Language Navigation Via Large Multimodal Models2024/11/1
- NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation2024/2/1
- GAMMA: Graspability-Aware Mobile MAnipulation Policy Learning based on Online Grasping Pose Fusion2023/9/1
- 3D-Aware Object Goal Navigation via Simultaneous Exploration and Identification2022/12/1
- GraspNeRF: Multiview-based 6-DoF Grasp Detection for Transparent and Specular Objects Using Generalizable NeRF2022/10/1