日本フィジカルAI新聞

世界のフィジカルAIを、日本語で。

週刊ニュースレター購読
arXiv:2606.01363v1

All Models are Wrong, Knowing Where is Useful: On Model Uncertainty in Reinforcement Learning

All Models are Wrong, Knowing Where is Useful: On Model Uncertainty in Reinforcement Learning

シェア:XThreadsFacebookLINEはてブBluesky

著者: Bernd Frauenknecht, Devdutt Subhasish, Artur Eisele, Friedrich Solowjow, Sebastian Trimpe

分類: cs.LG, eess.SY

原文アブストラクト

Model-based reinforcement learning (MBRL) infers information about the environment from a learned dynamics model and bears the potential to address open problems such as data efficient and safe learning in robotics. However, inaccuracies of the learned dynamics model are typically exploited by the agent, substantially hampering the capabilities of MBRL methods. We present a framework for dealing with inaccuracies of probabilistic models through targeted handling of uncertainty that effectively mitigates model exploitation. We present recent successes in learning directly on hardware and safe exploration, and discuss future directions for uncertainty-aware MBRL.