Rong Xiong
Zhejiang Humanoid Robot Innovation Center Co., Ltd.
収録論文 127本 ・ フィジカルAI/ロボット学習
ナビゲーション
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- 具現化ナビゲーター:指差し・思考・記憶・整合による効率的ナビゲーションナビゲーション2026/8/18
大規模視覚言語モデルを具現化ナビゲーションに活用する際の課題を解決するため、2Dピクセル選択と3D投影による行動形式、選択的推論と記憶圧縮、GRPOによる二段階整合を備えた統一フレームワークTAMP-Navを提案した。
- G2G: Exploiting Intra-Group Geometry for Inter-Group Pose Estimation2026/6/1
- APT: Action Expert Pretraining Improves Instruction Generalization of Vision-Language-Action Policies2026/6/1
- Learning Asynchronous Upper-body Task-space Trajectory Tracking Policy for Humanoid Robots2026/6/1
- Reference-Augmented Learning for Precise Tracking Policy of Tendon-Driven Continuum Robots2026/4/1
- {\Psi}-Map: Panoptic Surface Integrated Mapping Enables Real2Sim Transfer2026/4/1
- Learning-Based Dynamics Modeling and Robust Control for Tendon-Driven Continuum Robots2026/4/1
- Fast-SegSim: Real-Time Open-Vocabulary Segmentation for Robotics in Simulation2026/4/1
- IntentReact: Guiding Reactive Object-Centric Navigation via Topological Intent2026/3/1
- Direction Matters: Learning Force Direction Enables Sim-to-Real Contact-Rich Manipulation2026/2/1
- Seeing to Act, Prompting to Specify: A Bayesian Factorization of Vision Language Action Policy2025/12/1
- Neural Ranging Inertial Odometry2025/12/1
- Mr. Virgil: Learning Multi-robot Visual-range Relative Localization2025/12/1
- ETP-R1: Evolving Topological Planning with Reinforcement Fine-tuning for Vision-Language Navigation in Continuous Environments2025/12/1
- ZJUNlict Extended Team Description Paper 20252025/11/1
- High-Precision and High-Efficiency Trajectory Tracking for Excavators Based on Closed-Loop Dynamics2025/9/1
- BEV-ODOM2: Enhanced BEV-based Monocular Visual Odometry with PV-BEV Fusion and Dense Flow Supervision for Ground Robots2025/9/1
- Toward Embodiment Equivariant Vision-Language-Action Policy2025/9/1
- ExploreVLM: Closed-Loop Robot Exploration Task Planning with Vision-Language Models2025/8/1
- A Whole-Body Motion Imitation Framework from Human Data for Full-Size Humanoid Robot2025/8/1
- TOP: Time Optimization Policy for Stable and Accurate Standing Manipulation with Humanoid Robots2025/8/1
- EMP: Executable Motion Prior for Humanoid Robot Standing Upper-body Motion Imitation2025/7/1
- Grounding 3D Object Affordance with Language Instructions, Visual Observations and Interactions2025/4/1
- UnIRe: Unsupervised Instance Decomposition for Dynamic Urban Scene Reconstruction2025/4/1
- Natural Humanoid Robot Locomotion with Generative Motion Prior2025/3/12
- Natural Humanoid Robot Locomotion with Generative Motion Prior2025/3/1
- Efficient Alignment of Unconditioned Action Prior for Language-conditioned Pick and Place in Clutter2025/3/1
- Disambiguate Gripper State in Grasp-Based Tasks: Pseudo-Tactile as Feedback Enables Pure Simulation Learning2025/3/1
- PanopticSplatting: End-to-End Panoptic Gaussian Splatting2025/3/1
- CNSv2: Probabilistic Correspondence Encoded Neural Image Servo2025/3/1
- Compliance while resisting: a shear-thickening fluid controller for physical human-robot interaction2025/2/1
- BEV-DWPVO: BEV-based Differentiable Weighted Procrustes for Low Scale-drift Monocular Visual Odometry on Ground2025/2/1
- CarPlanner: Consistent Auto-regressive Trajectory Planning for Large-scale Reinforcement Learning in Autonomous Driving2025/2/1
- UniMM: A Unified Mixture Model Framework for Multi-Agent Simulation2025/1/1
- Leverage Cross-Attention for End-to-End Open-Vocabulary Panoptic Reconstruction2025/1/1
- Multi-cam Multi-map Visual Inertial Localization: System, Validation and Dataset2024/12/1
- BEV-ODOM: Reducing Scale Drift in Monocular Visual Odometry with BEV Representation2024/11/1
- LI-GS: Gaussian Splatting with LiDAR Incorporated for Accurate Large-Scale Reconstruction2024/9/1
- RING#: PR-by-PE Global Localization with Roto-translation Equivariant Gram Learning2024/9/1
- PanopticRecon: Leverage Open-vocabulary Instance Segmentation for Zero-shot Panoptic Reconstruction2024/7/1
- Pretraining-finetuning Framework for Efficient Co-design: A Case Study on Quadruped Robot Parkour2024/7/1
- $\nu$-DBA: Neural Implicit Dense Bundle Adjustment Enables Image-Only Driving Scene Reconstruction2024/4/1
- Grasp, See, and Place: Efficient Unknown Object Rearrangement with Policy Structure Prior2024/2/1
- Smooth Path Planning with Subharmonic Artificial Potential Field2024/2/1
- NGEL-SLAM: Neural Implicit Representation-based Global Consistent Low-Latency SLAM System2023/11/1
- A Two-stage Based Social Preference Recognition in Multi-Agent Autonomous Driving System2023/10/1
- CNS: Correspondence Encoded Neural Image Servo Policy2023/9/1
- Sparse Waypoint Validity Checking for Self-Entanglement-Free Tethered Path Planning2023/8/1
- 3D Model-free Visual Localization System from Essential Matrix under Local Planar Motion2023/8/1
- Leveraging BEV Representation for 360-degree Visual Place Recognition2023/5/1
- An Efficient Multi-solution Solver for the Inverse Kinematics of 3-Section Constant-Curvature Robots2023/5/1
- Object-centric Inference for Language Conditioned Placement: A Foundation Model based Approach2023/4/1
- A Hyper-network Based End-to-end Visual Servoing with Arbitrary Desired Poses2023/4/1
- Zero-shot Transfer Learning of Driving Policy via Socially Adversarial Traffic Flow2023/4/1
- Learning adaptive manipulation of objects with revolute joint: A case study on varied cabinet doors opening2023/4/1
- NF-Atlas: Multi-Volume Neural Feature Fields for Large Scale LiDAR Mapping2023/4/1
- GOOD: General Optimization-based Fusion for 3D Object Detection via LiDAR-Camera Object Candidates2023/3/1
- DAMS-LIO: A Degeneration-Aware and Modular Sensor-Fusion LiDAR-inertial Odometry2023/2/1
- Failure-aware Policy Learning for Self-assessable Robotics Tasks2023/2/1
- A Joint Modeling of Vision-Language-Action for Target-oriented Grasping in Clutter2023/2/1
- A Survey on Global LiDAR Localization: Challenges, Advances and Open Problems2023/2/1
- EMV-LIO: An Efficient Multiple Vision aided LiDAR-Inertial Odometry2023/2/1
- Open-Set Object Detection Using Classification-free Object Proposal and Instance-level Contrastive Learning2022/11/1
- RING++: Roto-translation Invariant Gram for Global Localization on a Sparse Scan Map2022/10/1
- DeepRING: Learning Roto-translation Invariant Representation for LiDAR based Place Recognition2022/10/1
- C^2:Co-design of Robots via Concurrent Networks Coupling Online and Offline Reinforcement Learning2022/9/1
- Efficient Distance-Optimal Tethered Path Planning in Planar Environments: The Workspace Convexity2022/8/1
- Towards Two-view 6D Object Pose Estimation: A Comparative Study on Fusion Strategy2022/7/1
- FEJ-VIRO: A Consistent First-Estimate Jacobian Visual-Inertial-Ranging Odometry2022/7/1
- Efficient Search of the k Shortest Non-Homotopic Paths by Eliminating Non-k-Optimal Topologies2022/7/1
- DPCN++: Differentiable Phase Correlation Network for Versatile Pose Registration2022/6/1
- Learning A Simulation-based Visual Policy for Real-world Peg In Unseen Holes2022/5/1
- Learning to Fill the Seam by Vision: Sub-millimeter Peg-in-hole on Unseen Shapes in Real World2022/4/1
- One RING to Rule Them All: Radon Sinogram for Place Recognition, Orientation and Translation Estimation2022/4/1
- Toward Consistent and Efficient Map-based Visual-inertial Localization: Theory Framework and Filter Design2022/4/1
- Map-based Visual-Inertial Localization: Consistency and Complexity2022/4/1
- Depth-Independent Depth Completion via Least Square Estimation2022/3/1
- DXQ-Net: Differentiable LiDAR-Camera Extrinsic Calibration Using Quality-aware Flow2022/3/1
- A Visual Navigation Perspective for Category-Level Object Pose Estimation2022/3/1
- Translation Invariant Global Estimation of Heading Angle Using Sinogram of LiDAR Point Cloud2022/3/1
- Electric Vehicle Automatic Charging System Based on Vision-force Fusion2021/10/1
- Learning Observation-Based Certifiable Safe Policy for Decentralized Multi-Robot Navigation2021/9/1
- Learning Interpretable BEV Based VIO without Deep Neural Networks2021/9/1
- Socially-Aware Multi-Agent Following with 2D Laser Scans via Deep Reinforcement Learning and Potential Field2021/9/1
- Efficient Object Manipulation to an Arbitrary Goal Pose: Learning-based Anytime Prioritized Planning2021/9/1
- Anti-degenerated UWB-LiDAR Localization for Automatic Road Roller in Tunnel2021/9/1
- Domain Generalization for Vision-based Driving Trajectory Generation2021/9/1
- HiTMap: A Hierarchical Topological Map Representation for Navigation in Unknown Environments2021/9/1
- Improved Radar Localization on Lidar Maps Using Shared Embedding2021/6/1
- Neural Motion Prediction for In-flight Uneven Object Catching2021/3/1
- Kinematic Motion Retargeting via Neural Latent Optimization for Learning Sign Language2021/3/1
- Efficient learning of goal-oriented push-grasping synergy in clutter2021/3/1
- Collaborative Recognition of Feasible Region with Aerial and Ground Robots through DPCN2021/3/1
- Toward Consistent Drift-free Visual Inertial Localization on Keyframe Based Map2021/3/1
- Learn to Differ: Sim2Real Small Defection Segmentation Network2021/3/1
- Radar-to-Lidar: Heterogeneous Place Recognition via Joint Learning2021/2/1
- PREGAN: Pose Randomization and Estimation for Weakly Paired Image Style Translation2020/11/1
- Improving Redundancy Availability: Dynamic Subtasks Modulation for Robots with Redundancy Insufficiency2020/11/1
- CORAL: Colored structural representation for bi-modal place recognition2020/11/1
- Dynamic Movement Primitive based Motion Retargeting for Dual-Arm Sign Language Motions2020/11/1
- Learning World Transition Model for Socially Aware Robot Navigation2020/11/1
- Robust localization for planar moving robot in changing environment: A perspective on density of correspondence and depth2020/11/1
- Improving the generalization of network based relative pose regression: dimension reduction as a regularizer2020/10/1
- DiSCO: Differentiable Scan Context with Orientation2020/10/1
- Imitation Learning of Hierarchical Driving Model: from Continuous Intention to Continuous Trajectory2020/10/1
- Deep Samplable Observation Model for Global Localization and Kidnapping2020/9/1
- RaLL: End-to-end Radar Localization on Lidar Map Using Differentiable Measurement Model2020/9/1
- Deep Phase Correlation for End-to-End Heterogeneous Sensor Measurements Matching2020/8/1
- Collaborative Localization of Aerial and Ground Mobile Robots through Orthomosaic Map2020/7/1
- Radar-on-Lidar: metric radar localization on prior lidar maps2020/5/1
- Learning hierarchical behavior and motion planning for autonomous driving2020/5/1
- Globally optimal consensus maximization for robust visual inertial localization in point and line map2020/2/1
- Cellular Decomposition for Non-repetitive Coverage Task with Minimum Discontinuities2020/1/1
- Weakly-Supervised Road Affordances Inference and Learning in Scenes without Traffic Signs2019/11/1
- DeepGoal: Learning to Drive with driving intention from Human Control Demonstration2019/11/1
- Champion Team Paper: Dynamic Passing-Shooting Algorithm Based on CUDA of The RoboCup SSL 2019 Champion2019/9/1
- Multi-agent Collaboration for Feasible Collaborative Behavior Construction and Evaluation2019/9/1
- Towards navigation without precise localization: Weakly supervised learning of goal-directed navigation cost map2019/6/1
- Efficient two step optimization for large embedded deformation graph based SLAM2019/6/1
- ZJUNlict Extended Team Description Paper for RoboCup 20192019/5/1
- Mechatronic Design of a Dribbling System for RoboCup Small Size Robot2019/5/1
- LiDAR-Camera Calibration under Arbitrary Configurations: Observability and Methods2019/3/1
- 2-Entity RANSAC for robust visual localization in changing environment2019/3/1
- Communication constrained cloud-based long-term visual localization in real time2019/3/1
- Multi-session Map Construction in Outdoor Dynamic Environment2018/7/1
- Laser map aided visual inertial localization in changing environment2018/3/1
- LocNet: Global localization in 3D point clouds for mobile vehicles2017/12/1