Nan Sun
Tsinghua University
収録論文 11本 ・ フィジカルAI/ロボット学習
VLA
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- 4D-WAM: 軌跡フィールドによる世界行動モデルへの時空間認識の注入VLA2026/8/8
ロボットの行動生成と動画予測を統合する世界行動モデルに、3次元軌跡フィールドの時空間知識を表現整合で注入する訓練戦略を提案。局所的な動き整合と長期的な目的地整合の2つの目的関数により、軌跡レベルの時空間表現を学習し、空間理解や実行精度、汎化性を向上させる。
- 4D-WAM: 軌跡フィールドによる世界行動モデルへの時空間認識の注入VLA2026/8/8
ロボットの行動生成と動画予測を統合する世界行動モデルに、3次元軌跡フィールドの時空間知識を表現整合で注入する訓練戦略を提案し、空間理解と実行精度を向上させた。
- 4D-WAM: Infusing Spatiotemporal Awareness into World Action Models through Trajectory Fields2026/8/1
- Xiaomi-Robotics-1: Scaling Vision-Language-Action Models with over 100K Hours of Real-World Trajectories2026/7/1
- Xiaomi-Robotics-U0: Unified Embodied Synthesis with World Foundation Model2026/7/1
- Revisiting Embodied Chain-of-Thought for Generalizable Robot Manipulation2026/6/1
- SpatialVAM:Spatial-Aware Multi-View Video Diffusion as a Data-Efficient Robot Policy2026/4/1
- Unified 4D World Action Modeling from Video Priors with Asynchronous Denoising2026/4/1
- Transforming Monolithic Foundation Models into Embodied Multi-Agent Architectures for Human-Robot Collaboration2025/12/1
- CollabVLA: Self-Reflective Vision-Language-Action Model Dreaming Together with Human2025/9/1
- AssistantX: An LLM-Powered Proactive Assistant in Collaborative Human-Populated Environment2024/9/1