Qi Wu
収録論文 44本 ・ フィジカルAI/ロボット学習
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- HyperDCM: Dynamic Cluster Memory Replay in Hyperbolic Space for Continual Robotic Navigation Across Scenes2026/7/1
- Embodied Agents Take Control: Minimal-Interface Zero-Shot Agents Rival Industrial-Scale Policies in Vision-and-Language Navigation2026/7/1
- From Region Arrival to Instance-Level Grounding in Vision-and-Language Navigation2026/7/1
- NVIDIA OmniDreams: Real-Time Generative World Model for Closed-Loop Autonomous Vehicle Simulation2026/6/1
- Automating the Design of Embodied Agent Architectures2026/6/1
- Ask When It Pays: Cost-Aware Open-Ended Interaction for Instance Goal Navigation2026/6/1
- SpatialAnt: Autonomous Zero-Shot Robot Navigation via Active Scene Reconstruction and Visual Anticipation2026/3/1
- Does Peer Observation Help? Vision-Sharing Collaboration for Vision-Language Navigation2026/3/1
- One Agent to Guide Them All: Empowering MLLMs for Vision-and-Language Navigation via Explicit World Representation2026/2/1
- SpatialNav: Leveraging Spatial Scene Graphs for Zero-Shot Vision-and-Language Navigation2026/1/1
- VLN-MME: Diagnosing MLLMs as Language-guided Visual Navigation agents2025/12/1
- VLNVerse: A Benchmark for Vision-Language Navigation with Versatile, Embodied, Realistic Simulation and Evaluation2025/12/1
- Arcadia: Toward a Full-Lifecycle Framework for Embodied Lifelong Learning2025/12/1
- Fast-SmartWay: Panoramic-Free End-to-End Zero-Shot Vision-and-Language Navigation2025/11/1
- SimULi: Real-Time LiDAR and Camera Simulation with Unscented Transforms2025/10/1
- Embodied Navigation Foundation Model2025/9/1
- Recursive Visual Imagination and Adaptive Linguistic Grounding for Vision Language Navigation2025/7/1
- NeuroLoc: Encoding Navigation Cells for 6-DOF Camera Localization2025/5/1
- BadNAVer: Exploring Jailbreak Attacks On Vision-and-Language Navigation2025/5/1
- Dribble Master: Learning Agile Humanoid Dribbling through Legged Locomotion2025/5/1
- COSMO: Combination of Selective Memorization for Low-cost Vision-and-Language Navigation2025/3/1
- SmartWay: Enhanced Waypoint Prediction and Backtracking for Zero-Shot Vision-and-Language Navigation2025/3/1
- Navigating Motion Agents in Dynamic and Cluttered Environments through LLM Reasoning2025/3/1
- Ground-level Viewpoint Vision-and-Language Navigation in Continuous Environments2025/2/1
- SAME: Learning Generic Language-Guided Visual Navigation with State-Adaptive Mixture of Experts2024/12/1
- Helpful DoggyBot: Open-World Object Fetching using Legged Robots and Vision-Language Models2024/10/1
- Open-Nav: Exploring Zero-Shot Vision-and-Language Navigation in Continuous Environment with Open-Source LLMs2024/9/1
- NavGPT-2: Unleashing Navigational Reasoning Capability for Large Vision-Language Models2024/7/1
- Navigating Beyond Instructions: Vision-and-Language Navigation in Obstructed Environments2024/7/1
- HumanPlus: Humanoid Shadowing and Imitation from Humans2024/6/1
- TON-VIO: Online Time Offset Modeling Networks for Robust Temporal Alignment in High Dynamic Motion VIO2024/3/1
- Thermal-NeRF: Neural Radiance Fields from an Infrared Camera2024/3/1
- Explicit Interaction for Fusion-Based Place Recognition2024/2/1
- NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation2024/2/1
- AerialVLN: Vision-and-Language Navigation for UAVs2023/8/1
- NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models2023/5/1
- Custom Sine Waves Are Enough for Imitation Learning of Bipedal Gaits with Different Styles2022/4/1
- Bridging the Gap Between Learning in Discrete and Continuous Environments for Vision-and-Language Navigation2022/3/1
- Adaptive Mimic: Deep Reinforcement Learning of Parameterized Bipedal Walking from Infeasible References2021/12/1
- Communicative Learning with Natural Gestures for Embodied Navigation Agents with Human-in-the-Scene2021/8/1
- Semantics for Robotic Mapping, Perception and Interaction: A Survey2021/1/1
- P3-LOAM: PPP/LiDAR Loosely Coupled SLAM with Accurate Covariance Estimation and Robust RAIM in Urban Canyon Environment2020/12/1
- Attention-SLAM: A Visual Monocular SLAM Learning from Human Gaze2020/9/1
- Vision-and-Language Navigation: Interpreting visually-grounded navigation instructions in real environments2017/11/1