Yi Liu
Shanghai Innovation Institute
収録論文 29本 ・ フィジカルAI/ロボット学習
VLA
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- τ0-VLA: 世界モデル誘導のテスト時計算を備えた階層型ロボット基盤モデルVLA2026/8/17
長期的なロボット操作タスクを、高レベル方策がサブタスク生成時にテスト時計算を追加で行える階層型VLAモデルを提案し、実データで性能向上を実証した。
- STAR-VLM: Spatiotemporal Grounding Vision-Language Models for Motion and Velocity Estimation via Automotive Radar Supervision2026/8/1
- Learning Spatiotemporal Decision Priors for Efficient Path Planning under Partial Observability2026/7/1
- CRISP: A Spatiotemporal Camera-Radar Backbone for Driving via Forecasting-Based World-Model Pretraining2026/7/1
- HCSG: Human-Centric Semantic-Geometric Reasoning for Vision-Language Navigation2026/5/1
- Safety in Embodied AI: A Survey of Risks, Attacks, and Defenses2026/5/1
- GGD-SLAM: Monocular 3DGS SLAM Powered by Generalizable Motion Model for Dynamic Environments2026/4/1
- Libra-VLA: Achieving Learning Equilibrium via Asynchronous Coarse-to-Fine Dual-System2026/4/1
- RMGS-SLAM: Real-time Multi-sensor Gaussian Splatting SLAM2026/4/1
- HazardArena: Evaluating Semantic Safety in Vision-Language-Action Models2026/4/1
- OVAL: Open-Vocabulary Augmented Memory Model for Lifelong Object Goal Navigation2026/4/1
- TaPD: Temporal-adaptive Progressive Distillation for Observation-Adaptive Trajectory Forecasting in Autonomous Driving2026/3/1
- Recover to Predict: Progressive Retrospective Learning for Variable-Length Trajectory Prediction2026/3/1
- Enhancing Vision-Language Navigation with Multimodal Event Knowledge from Real-World Indoor Tour Videos2026/2/1
- SOP: A Scalable Online Post-Training System for Vision-Language-Action Models2026/1/1
- ACoT-VLA: Action Chain-of-Thought for Vision-Language-Action Models2026/1/1
- CausalNav: A Long-term Embodied Navigation System for Autonomous Mobile Robots in Dynamic Outdoor Scenarios2026/1/1
- Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training2025/12/1
- FishDetector-R1: Unified MLLM-Based Framework with Reinforcement Fine-Tuning for Weakly Supervised Fish Detection, Segmentation, and Counting2025/12/1
- DyPho-SLAM : Real-time Photorealistic SLAM in Dynamic Environments2025/9/1
- Benchmarking Generalizable Bimanual Manipulation: RoboTwin Dual-Arm Collaboration Challenge at CVPR 2025 MEIS Workshop2025/6/1
- AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems2025/3/1
- Range-SLAM: Ultra-Wideband-Based Smoke-Resistant Real-Time Localization and Mapping2024/9/1
- ExploitFlow, cyber security exploitation routes for Game Theory and AI research in robotics2023/8/1
- DSL-Assembly: A Robust and Safe Assembly Strategy2023/2/1
- SIRL: Similarity-based Implicit Representation Learning2023/1/1
- GOMP-FIT: Grasp-Optimized Motion Planning for Fast Inertial Transport2021/10/1
- Real-Time Trajectory Planning for AGV in the Presence of Moving Obstacles: A First-Search-Then-Optimization Approach2019/2/1
- Parallax Bundle Adjustment on Manifold with Convexified Initialization2018/7/1