日本フィジカルAI新聞

世界のフィジカルAIを、日本語で。

週刊ニュースレター購読
VLAarXiv:2606.00145

境界での完了判定(CaB):限定的なキャリブレーション下での完了認識制御による展開可能な切り替え

Completion at the Boundary (CaB): Deployable Switching with Completion-Aware Control under Limited Calibration

シェア:XThreadsFacebookLINEはてブBluesky

VLAエージェントの複合指示実行において、指示完了のタイミングを判定する新しい手法CaBを提案。境界前後の両側情報を保持するトークンで切り替え判断と動作生成を安定化する。

著者: Yusuke Sano, Takeshi Itoga

分類: cs.RO, cs.AI

原文アブストラクト

Vision-language-action (VLA) agents can execute natural-language instructions, yet deployed systems still lack an operational interface: deciding when the instruction is complete. This gap is acute in short composites ("do A, then B"), where mistimed handoffs cascade into downstream failures. Completion is inherently closed-loop because switching is an intervention that changes the instruction context and thus future actions and observations. We study completion under a deployable low-calibration regime motivated by open-ended instruction spaces, enforcing no test-time relearning and a single globally calibrated switching rule selected once on development set and reused unchanged on test set. Under this constraint, collapsing asymmetric boundary evidence into a single scalar can be brittle under polarity shifts across tasks. We propose Completion at the Boundary (CaB), which predicts an event-local completion object in the form of Boundary-Phase Tokens (Before/Hit/After), retaining two-sided boundary evidence under this discipline. CaB-When converts this completion object into a minimal, auditable switching decision (when), while CaB-How reuses the same completion object to condition action generation for boundary-stable control through handoffs (how). Using an intervention-aware E1/E2 protocol, we show that CaB improves composite execution and handoff quality on a first-person Minecraft VLA benchmark under matched capacity and deployability constraints.

関連論文