Jiabing Yang
Institute of Automation, Chinese Academy of Sciences
収録論文 13本 ・ フィジカルAI/ロボット学習
ワールドモデルVLA/操作
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- XEWorld:行動条件付きワールドモデルは未知のロボット形態に汎化できるか?ワールドモデル2026/8/6
ロボット操作のための行動条件付きワールドモデルが、未見のロボット形態に対して物理ダイナミクスを正しく予測できるかを検証するため、クロスエンボディメントテストベッドXEWorldを導入し、既存モデルの限界を分析した。
- BridgeVLA++: データ効率的で汎化性が高く、メモリ拡張された3D操作のための視覚-言語-行動フレームワークVLA/操作2026/8/5
事前学習済み視覚言語モデルを活用した3Dロボット操作フレームワークBridgeVLAを拡張し、空間的・時間的メモリを統合することで、データ効率と汎化性を保ちつつ、記憶依存の操作タスクで最先端の性能を達成した。
- XEWorld: Can Action-Conditioned World Models Generalize to Unseen Robot Embodiments?2026/8/1
- BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation2026/8/1
- FlowWAM: Optical Flow as a Unified Action Representation for World Action Models2026/7/1
- Improving Vision-Language-Action Model Fine-Tuning with Structured Stage and Keyframe Supervision2026/6/1
- DIM-WAM: World-Action Modeling with Diverse Historical Event Memory2026/6/1
- SKIP: Sparse Keyframe Interpolation Paradigm for Efficient Embodied World Models2026/6/1
- SpatialVAM:Spatial-Aware Multi-View Video Diffusion as a Data-Efficient Robot Policy2026/4/1
- UAOR: Uncertainty-aware Observation Reinjection for Vision-Language-Action Models2026/2/1
- BridgeV2W: Bridging Video Generation Models to Embodied World Models via Embodiment Masks2026/2/1
- EgoDemoGen: Egocentric Demonstration Generation for Viewpoint Generalization in Robotic Manipulation2025/9/1
- EC-Flow: Enabling Versatile Robotic Manipulation from Action-Unlabeled Videos via Embodiment-Centric Flow2025/7/1