Hongsheng Li
収録論文 31本 ・ フィジカルAI/ロボット学習
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- ACE-Ego-0: Unifying Egocentric Human and Robotic Data for VLA Pretraining2026/6/1
- OneVLA: A Unified Framework for Embodied Tasks2026/6/1
- Hierarchical Advantage Weighting for Online RL Fine-Tuning of VLAs from Sparse Episode Outcomes2026/6/1
- MindVLA-U1: VLA Beats VA with Unified Streaming Architecture for Autonomous Driving2026/5/1
- DexJoCo: A Benchmark and Toolkit for Task-Oriented Dexterous Manipulation on MuJoCo2026/5/1
- LMGenDrive: Bridging Multimodal Understanding and Generative World Modeling for End-to-End Driving2026/4/1
- Walk With Me: Long-Horizon Social Navigation for Human-Centric Outdoor Assistance2026/4/1
- DrivingGen: A Comprehensive Benchmark for Generative Video World Models in Autonomous Driving2026/1/1
- EnerVerse-AC: Envisioning Embodied Environments with Action Condition2025/5/1
- UAV-Flow Colosseo: A Real-World Benchmark for Flying-on-a-Word UAV Imitation Learning2025/5/1
- Adversarial Data Collection: Human-Collaborative Perturbations for Efficient and Robust Robotic Imitation Learning2025/3/1
- EnerVerse: Envisioning Embodied Future Space for Robotics Manipulation2025/1/1
- ZOPP: A Framework of Zero-shot Offboard Panoptic Perception for Autonomous Driving2024/11/1
- Towards Realistic UAV Vision-Language Navigation: Platform, Benchmark, and Methodology2024/10/1
- SmartPretrain: Model-Agnostic and Dataset-Agnostic Representation Learning for Motion Prediction2024/10/1
- UniAff: A Unified Representation of Affordances for Tool Usage and Articulation with Vision-Language Models2024/9/1
- SKT: Integrating State-Aware Keypoint Trajectories with Vision-Language Models for Robotic Garment Manipulation2024/9/1
- A3VLM: Actionable Articulation-Aware Vision Language Model2024/6/1
- SmartRefine: A Scenario-Adaptive Refinement Framework for Efficient Motion Prediction2024/3/1
- ManipVQA: Injecting Robotic Affordance and Physically Grounded Information into Multi-Modal Large Language Models2024/3/1
- LMDrive: Closed-Loop End-to-End Driving with Large Language Models2023/12/1
- Instruct2Act: Mapping Multi-modality Instructions to Robotic Actions with Large Language Model2023/5/1
- Perception Imitation: Towards Synthesis-free Simulator for Autonomous Vehicles2023/4/1
- ConQueR: Query Contrast Voxel-DETR for 3D Object Detection2022/12/1
- Safety-Enhanced Autonomous Driving Using Interpretable Sensor Fusion Transformer2022/7/1
- 3D Object Detection for Autonomous Driving: A Comprehensive Survey2022/6/1
- RNNPose: Recurrent 6-DoF Object Pose Refinement with Robust Correspondence Field Estimation and Pose Optimization2022/3/1
- Robust Self-Supervised LiDAR Odometry via Representative Structure Discovery and 3D Inherent Error Modeling2022/2/1
- PointCLIP: Point Cloud Understanding by CLIP2021/12/1
- Progressive Correspondence Pruning by Consensus Learning2021/1/1
- SelfVoxeLO: Self-supervised LiDAR Odometry with Voxel-based Deep Neural Networks2020/10/1