Lingdong Kong
収録論文 59本 ・ フィジカルAI/ロボット学習
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- Quo Vadis, World Modeling?2026/8/1
- Data Pyramid for Embodied Manipulation: A Survey2026/7/1
- Token-Wise Latent Streaming from Slow Reasoners to Fast Planners for Dynamic Vision Language Navigation2026/7/1
- Not All Points Are Equal: Uncertainty-Aware 4D LiDAR Scene Synthesis2026/6/1
- Is Your Driving World Model an All-Around Player?2026/5/1
- OmniLiDAR: A Unified Diffusion Framework for Multi-Domain 3D LiDAR Generation2026/5/1
- Xiaomi OneVL: One-Step Latent Reasoning and Planning with Vision-Language Explanation2026/4/1
- Language-Conditioned World Modeling for Visual Navigation2026/3/1
- NavThinker: Action-Conditioned World Models for Coupled Prediction and Planning in Social Navigation2026/3/1
- FLUX: Accelerating Cross-Embodiment Generative Navigation Policies via Rectified Flow and Static-to-Dynamic Learning2026/3/1
- The RoboSense Challenge: Sense Anything, Navigate Anywhere, Adapt Across Platforms2026/1/1
- Vision-Language-Action Models for Autonomous Driving: Past, Present, and Future2025/12/1
- U4D: Uncertainty-Aware 4D World Modeling from LiDAR Sequences2025/12/1
- Forging Spatial Intelligence: A Roadmap of Multi-Modal Data Pre-Training for Autonomous Systems2025/12/1
- 3EED: Ground Everything Everywhere in 3D2025/11/1
- Towards Unified World Models for Visual Navigation via Memory-Augmented Planning and Foresight2025/10/1
- 3D and 4D World Modeling: A Survey2025/9/1
- Learning to Generate 4D LiDAR Sequences2025/9/1
- Visual Grounding from Event Cameras2025/9/1
- La La LiDAR: Large-Scale Layout Generation from LiDAR Data2025/8/1
- LiDARCrafter: Dynamic 4D World Modeling from LiDAR Sequences2025/8/1
- Veila: Panoramic LiDAR Generation from a Monocular RGB Image2025/8/1
- Beyond One Shot, Beyond One Perspective: Cross-View and Long-Horizon Distillation for Better LiDAR Representations2025/7/1
- Monocular Semantic Scene Completion via Masked Recurrent Networks2025/7/1
- Perspective-Invariant 3D Object Detection2025/7/1
- Talk2Event: Grounded Understanding of Dynamic Scenes from Event Cameras2025/7/1
- Zero-Shot 3D Visual Grounding from Vision-Language Models2025/5/1
- Stairway to Success: An Online Floor-Aware Zero-Shot Object-Goal Navigation Framework via LLM-Driven Coarse-to-Fine Exploration2025/5/1
- HA-VLN 2.0: An Open Benchmark and Leaderboard for Human-Aware Navigation in Discrete and Continuous Environments with Dynamic Multi-Human Interactions2025/3/1
- Enhanced Spatiotemporal Consistency for Image-to-LiDAR Data Pretraining2025/3/1
- EventFly: Event Camera Perception from Ground to the Sky2025/3/1
- LargeAD: Large-Scale Cross-Sensor Data Pretraining for Autonomous Driving2025/1/1
- LiMoE: Mixture of LiDAR Representation Learners from Automotive Scenes2025/1/1
- Are VLMs Ready for Autonomous Driving? An Empirical Study from the Reliability, Data, and Metric Perspectives2025/1/1
- MSC-Bench: Benchmarking and Analyzing Multi-Sensor Corruption for Driving Perception2025/1/1
- SeeGround: See and Ground for Zero-Shot Open-Vocabulary 3D Visual Grounding2024/12/1
- FlexEvent: Towards Flexible Event-Frame Object Detection at Varying Operational Frequencies2024/12/1
- DynamicCity: Large-Scale 4D Occupancy Generation from Dynamic Scenes2024/10/1
- 4D Contrastive Superflows are Dense 3D Representation Learners2024/7/1
- Is Your HD Map Constructor Reliable under Sensor Corruptions?2024/6/1
- An Empirical Study of Training State-of-the-Art LiDAR Segmentation Models2024/5/1
- Benchmarking and Improving Bird's Eye View Perception Robustness in Autonomous Driving2024/5/1
- The RoboDrive Challenge: Drive Anytime Anywhere in Any Condition2024/5/1
- Multi-Modal Data-Efficient 3D Scene Understanding for Autonomous Driving2024/5/1
- OpenESS: Event-based Semantic Scene Understanding with Open Vocabularies2024/5/1
- Multi-Space Alignments Towards Universal LiDAR Segmentation2024/5/1
- Calib3D: Calibrating Model Preferences for Reliable 3D Scene Understanding2024/3/1
- Is Your LiDAR Placement Optimized for 3D Scene Understanding?2024/3/1
- FRNet: Frustum-Range Networks for Scalable LiDAR Segmentation2023/12/1
- RoboDepth: Robust Out-of-Distribution Depth Estimation under Corruptions2023/10/1
- The RoboDepth Challenge: Methods and Advancements Towards Robust Depth Estimation2023/7/1
- Segment Any Point Cloud Sequences by Distilling Vision Foundation Models2023/6/1
- RoboBEV: Towards Robust Bird's Eye View Perception under Corruptions2023/4/1
- Rethinking Range View Representation for LiDAR Segmentation2023/3/1
- Robo3D: Towards Robust and Reliable 3D Perception against Corruptions2023/3/1
- LaserMix for Semi-Supervised LiDAR Semantic Segmentation2022/7/1
- ConDA: Unsupervised Domain Adaptation for LiDAR Segmentation via Regularized Domain Concatenation2021/11/1
- Modification of Gesture-Determined-Dynamic Function with Consideration of Margins for Motion Planning of Humanoid Robots2020/8/1
- Kinematic Resolutions of Redundant Robot Manipulators using Integration-Enhanced RNNs2020/8/1