遠くへ届くには近くを狙え:凍結世界モデルは思うより賢く計画できる
Aim Short to Reach Far: Your Frozen World Model Can Plan Better Than You Think
凍結した視覚世界モデルでも、最終ゴールではなく経験から取得した中間目標を狙うことで、長距離タスクの計画性能が大幅に向上することを示した研究。
詳しい要約
1. どんなもの?
2. 先行研究と比べてどこがすごい?
3. 技術・手法の肝は?
4. どうやって有効だと検証した?
5. 議論はある?
6. 次に読むべき論文は?
※ AIが要旨から生成した要約です。正確性は原文をご確認ください。
著者: Xvyuan Liu, Jianjie Fang, Chen Gao, Yong Li
分類: cs.LG, cs.RO
原文アブストラクト
Planners built on visual world models commonly score each predicted outcome by its distance to the encoded goal image. We show that this target can limit control even with exact dynamics and globally optimal short-horizon search: reaching a goal may require actions that initially move away from it. With frozen LeWM models, intermediate targets substantially improve action synthesis and recorded-action ranking on Cube, PushT, Reacher, and TwoRoom. Learned targets and targets drawn from observed experience both produce these gains. We introduce Anchored Planning, which retrieves a recorded segment whose start and end resemble the current and goal observations, then aims at an observation shortly after its start. The frozen model scores actions toward this target from the current state. Without additional training, planning toward observed targets outperforms the released LeWM planner on every task in our long-range evaluation. Additional final-goal search falls short of the same gains. Lower successor-prediction error need not translate into better control. Success also depends on how far ahead the target is placed and on shrinking the retrieval span as execution advances. Changing only the target lets the same frozen model and planner reach goals that final-goal scoring misses.
関連論文
- プリズマティック・ワールドモデル:ハイブリッド系の計画のための合成可能なダイナミクス学習モデルベース計画
- ロバスト計画のための因果構造分布の学習モデルベース計画