日本フィジカルAI新聞

世界のフィジカルAIを、日本語で。

週刊ニュースレター購読
arXiv:2505.12350

Multi-CALF: A Policy Combination Approach with Statistical Guarantees

Multi-CALF: A Policy Combination Approach with Statistical Guarantees

シェア:XThreadsFacebookLINEはてブBluesky

著者: Georgiy Malaniya, Anton Bolychev, Grigory Yaremenko, Anastasia Krasnaya, Pavel Osinenko

分類: cs.LG, cs.AI, cs.RO, cs.SY, eess.SY, math.OC

原文アブストラクト

We introduce Multi-CALF, an algorithm that intelligently combines reinforcement learning policies based on their relative value improvements. Our approach integrates a standard RL policy with a theoretically-backed alternative policy, inheriting formal stability guarantees while often achieving better performance than either policy individually. We prove that our combined policy converges to a specified goal set with known probability and provide precise bounds on maximum deviation and convergence time. Empirical validation on control tasks demonstrates enhanced performance while maintaining stability guarantees.