Xiao Liu
Tsinghua University
収録論文 36本 ・ フィジカルAI/ロボット学習
所属歴(論文より): Tsinghua University(〜2026) / Harbin Institute of Technology(〜2026)
VLA操作マニピュレーションナビゲーション
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- G0.5: ロボットの推論と行動のための単一自己回帰ストリームVLA2026/8/12
事前学習済みのVLMと別の行動エキスパートを組み合わせる従来のVLAモデルに対し、単一のトランスフォーマーデコーダが推論トークンと行動トークンを単一の目的で生成する自己回帰VLAモデルG0.5を提案。クロスエンボディメント行動トークナイザー、ネイティブな思考連鎖ストリーム、視覚メモリモジュールにより、基礎モデル規模での学習を可能にし、複数のベンチマークで最先端を達成。
- StageWAM: ロボット操作におけるワールドアクションモデルのためのジョイント埋め込みステージ予測操作2026/8/11
ロボット操作タスクにおいて、短期的な物理的未来に加えて、タスクの進行段階を表すセマンティックな未来を予測するStageWAMを提案し、成功率と実行効率を向上させた。
- JEPA-WAM: ロボット操作のためのワールドアクションモデルにおける段階レベル結合埋め込み予測マニピュレーション2026/8/11
ロボット操作タスクにおいて、短期的な物理的未来と段階的な意味的未来を区別し、段階レベルの潜在目標を予測するJEPA-WAMを提案。50のタスクで成功率90.25%を達成し、実行ステップ数を削減した。
- SAIN: 能動的対話による構造認識型インタラクティブナビゲーションナビゲーション2026/8/10
曖昧な指示に対して能動的質問で解消する移動ロボットのナビゲーション手法を提案。対話の回答を永続的な空間・物体記憶に変換し、ゼロショットで目標探索を改善した。
- JEPA-WAM: Stage-Level Joint-Embedding Prediction for World-Action Models in Robot Manipulation2026/8/1
- SAIN: Structure-Aware Interactive Navigation with Active Dialogue Grounding for Mobile Robot2026/8/1
- G0.5: One Autoregressive Stream for Robot Reasoning and Action2026/8/1
- Semantic-Guided Progressive Object Removal with Gaussian Splatting2026/7/1
- NavCMPO: Critic-Guided MeanFlow Policy Optimization for Adaptive Navigation2026/7/1
- SoftNav: Injecting 3D Scene Tokens into VLMs for Embodied Navigation2026/7/1
- SUBTA: A Framework for Supported User-Guided Bimanual Teleoperation in Structured Assembly2026/3/1
- High-Precision and High-Efficiency Trajectory Tracking for Excavators Based on Closed-Loop Dynamics2025/9/1
- Leg-Arm Coordinated Operation for Curtain Wall Installation2025/9/1
- Galaxea Open-World Dataset and G0 Dual-System VLA Model2025/9/1
- Multi-Objective Trajectory Planning for a Robotic Arm in Curtain Wall Installation2025/7/1
- Dynamic Modeling and Dimensional Optimization of Legged Mechanisms for Construction Robot2025/7/1
- Dynamic Parameter Identification of a Curtain Wall Installation Robotic Arm2025/7/1
- Topology Optimization of Leg Structures for Construction Robots Based on Variable Density Method2025/7/1
- Trajectory Planning of a Curtain Wall Installation Robot Based on Biomimetic Mechanisms2025/7/1
- Design and Dimensional Optimization of Legged Structures for Construction Robots2025/7/1
- Enabling Stateful Behaviors for Diffusion-based Policy Learning2024/4/1
- Safe Hybrid-Action Reinforcement Learning-Based Decision and Control for Discretionary Lane Change2024/3/1
- Autonomous vehicle decision and control through reinforcement learning with traffic flow randomization2024/3/1
- iRoCo: Intuitive Robot Control From Anywhere Using a Smartwatch2024/3/1
- Discretionary Lane-Change Decision and Control via Parameterized Soft Actor-Critic for Hybrid Action Space2024/2/1
- Multimodal Learning of Soft Robot Dynamics using Differentiable Filters2023/11/1
- Probabilistic Differentiable Filters Enable Ubiquitous Robot Control with Smartwatches2023/9/1
- Learning Soft Robot Dynamics using Differentiable Kalman Filters and Spatio-Temporal Embeddings2023/8/1
- Enhancing State Estimation in Robots: A Data-Driven Approach with Differentiable Ensemble Kalman Filters2023/8/1
- LOF: Structure-Aware Line Tracking based on Optical Flow2021/9/1
- Speed Planning Using Bezier Polynomials with Trapezoidal Corridors2021/4/1
- Improved Signed Distance Function for 2D Real-time SLAM and Accurate Localization2021/1/1
- Robotic Communications for 5G and Beyond: Challenges and Research Opportunities2020/12/1
- Leveraging Planar Regularities for Point Line Visual-Inertial Odometry2020/4/1
- Coarse-To-Fine Visual Localization Using Semantic Compact Map2019/10/1
- DF-SLAM: A Deep-Learning Enhanced Visual SLAM System based on Deep Local Features2019/1/1