Hong Zhang
収録論文 70本 ・ フィジカルAI/ロボット学習
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- Traj-VLN: Learning Pixel-Space Interaction via Autoregressive Trajectory Generation2026/7/1
- ReferTrack: Referring Then Tracking for Embodied Visual Tracking2026/7/1
- HAM-VLN: Harnessing Hierarchical Agentic Memory for Zero-Shot Vision-and-Language Navigation2026/7/1
- GCNGrasp-VP: Affordance-Guided View Planning for Efficient Task-Oriented Grasping2026/6/1
- FLM-Occ: Feed-forward Likelihood Maximization for Efficient Indoor Occupancy Prediction2026/6/1
- Can Single-View Mesh Reconstruction Generalize to Robot Camera Rotation?2026/6/1
- TARIC: Memory-Augmented Traversability-Aware Outdoor VLN under Interrupted Semantic Cues2026/5/1
- Can Aerial VLA Models Cooperate? Evaluating Closed-Loop Air-Ground Coordination with CARLA-Air2026/5/1
- VLAConf: Calibrated Task-Success Confidence for Vision-Language-Action Models2026/5/1
- Enhancing Glass Surface Reconstruction via Depth Prior for Robot Navigation2026/4/1
- ProDrive: Proactive Planning for Autonomous Driving via Ego-Environment Co-Evolution2026/4/1
- Grasp as You Dream: Imitating Functional Grasping from Generated Human Demonstrations2026/4/1
- Easy-IIL: Reducing Human Operational Burden in Interactive Imitation Learning via Assistant Experts2026/3/1
- Edge-Assisted Multi-Robot Visual-Inertial SLAM with Efficient Communication2026/3/1
- Agentic Self-Evolutionary Replanning for Embodied Navigation2026/3/1
- CARLA-Air: Fly Drones Inside a CARLA World -- A Unified Infrastructure for Air-Ground Embodied Intelligence2026/3/1
- MfNeuPAN: Proactive End-to-End Navigation in Dynamic Environments via Direct Multi-Frame Point Constraints2025/11/1
- TrackVLA++: Unleashing Reasoning and Memory Capabilities in VLA Models for Embodied Visual Tracking2025/10/1
- Unveiling Uncertainty-Aware Autonomous Cooperative Learning Based Planning Strategy2025/10/1
- VG-Mapping: Variation-aware Density Control for Online 3D Gaussian Mapping in Semi-static Scenes2025/10/1
- Adap-RPF: Adaptive Trajectory Sampling for Robot Person Following in Dynamic Crowded Environments2025/10/1
- EZREAL: Enhancing Zero-Shot Outdoor Robot Navigation toward Distant Targets under Varying Visibility2025/9/1
- Meta-Memory: Retrieving and Integrating Semantic-Spatial Memories for Robot Spatial Reasoning2025/9/1
- Follow-Bench: A Unified Motion Planning Benchmark for Socially-Aware Robot Person Following2025/9/1
- MimicFunc: Imitating Tool Manipulation from a Single Human Video via Functional Correspondence2025/8/1
- GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions2025/8/1
- ROVER: Robust Loop Closure Verification with Trajectory Prior in Repetitive Environments2025/8/1
- JAM: Keypoint-Guided Joint Prediction after Classification-Aware Marginal Proposal for Multi-Agent Interaction2025/7/1
- TPT-Bench: A Large-Scale, Long-Term and Robot-Egocentric Dataset for Benchmarking Target Person Tracking2025/5/1
- Dexterous Manipulation through Imitation Learning: A Survey2025/4/1
- FlowPlan: Zero-Shot Task Planning with LLM Flow Engineering for Robotic Instruction Following2025/3/1
- Monocular Person Localization under Camera Ego-motion2025/3/1
- TextInPlace: Indoor Visual Place Recognition in Repetitive Structures with Scene Text Spotting and Verification2025/3/1
- RPF-Search: Field-based Search for Robot Person Following in Unknown Dynamic Environments2025/3/1
- ReJSHand: Efficient Real-Time Hand Pose Estimation and Mesh Reconstruction Using Refined Joint and Skeleton Features2025/3/1
- HGDiffuser: Efficient Task-Oriented Grasp Generation via Human-Guided Grasp Diffusion Models2025/3/1
- FUNCTO: Function-Centric One-Shot Imitation Learning for Tool Manipulation2025/2/1
- Optimizing NeRF-based SLAM with Trajectory Smoothness Constraints2024/10/1
- RTAGrasp: Learning Task-Oriented Grasping from Human Videos via Retrieval, Transfer, and Alignment2024/9/1
- FLAF: Focal Line and Feature-constrained Active View Planning for Visual Teach and Repeat2024/9/1
- GV-Bench: Benchmarking Local Feature Matching for Geometric Verification of Long-term Loop Closure Detection2024/7/1
- Human Orientation Estimation under Partial Observation2024/4/1
- FoundationGrasp: Generalizable Task-Oriented Grasping with Foundation Models2024/4/1
- Commonsense Scene Graph-based Target Localization for Object Search2024/4/1
- Person Re-Identification for Robot Person Following with Online Continual Learning2023/9/1
- Efficient Object Rearrangement via Multi-view Fusion2023/9/1
- GraspGPT: Leveraging Semantic Knowledge from a Large Language Model for Task-Oriented Grasping2023/7/1
- Prediction of SLAM ATE Using an Ensemble Learning Regression Model and 1-D Global Pooling of Data Characterization2023/3/1
- Task-Oriented Grasp Prediction with Visual-Language Inputs2023/2/1
- Robot Person Following Under Partial Occlusion2023/2/1
- NDD: A 3D Point Cloud Descriptor Based on Normal Distribution for Loop Closure Detection2022/9/1
- Optimizing SLAM Evaluation Footprint Through Dynamic Range Coverage Analysis of Datasets2022/9/1
- Following Closely: A Robust Monocular Person Following System for Mobile Robot2022/4/1
- Condition-Invariant and Compact Visual Place Description by Convolutional Autoencoder2022/4/1
- Mapping While Following: 2D LiDAR SLAM in Indoor Dynamic Environments with a Person Tracker2022/4/1
- Are We Ready for Robust and Resilient SLAM? A Framework For Quantitative Characterization of SLAM Datasets2022/2/1
- Online Mutual Adaptation of Deep Depth Prediction and Visual SLAM2021/11/1
- Relationship Oriented Affordance Learning through Manipulation Graph Construction2021/10/1
- Curiosity-based Robot Navigation under Uncertainty in Crowded Environments2021/6/1
- Polarimetric Monocular Dense Mapping Using Relative Deep Depth Prior2021/2/1
- PL-VINS: Real-Time Monocular Visual-Inertial SLAM with Point and Line Features2020/9/1
- Fast ORB-SLAM without Keypoint Descriptors2020/8/1
- DeepRelativeFusion: Dense Monocular SLAM using Single-Image Relative Depth Prediction2020/6/1
- Semi-Supervised Monocular Depth Estimation with Left-Right Consistency Using Deep Neural Network2019/5/1
- Towards A Deep Insight into Landmark-based Visual Place Recognition: Methodology and Practice2018/8/1
- Submap-based Pose-graph Visual SLAM: A Robust Visual Exploration and Localization System2018/7/1
- COROLA: A Sequential Solution to Moving Object Detection Using Low-rank Approximation2015/5/1
- Convolutional Neural Network-Based Image Representation for Visual Loop Closure Detection2015/4/1
- Stopping Rules for Bag-of-Words Image Search and Its Application in Appearance-Based Localization2013/12/1
- An Efficient Index for Visual Search in Appearance-based SLAM2013/9/1