シミュレーションから実世界への性能証明書のためのベッティング
Betting for Sim-to-Real Performance Certificates
実世界の試行結果を予測するためにシミュレーション結果を賭けとして利用し、実世界の結果に基づいて賭け金を増減させることで、任意の時点で有効な性能証明書(平均値の信頼区間)を生成するフレームワークを提案した。
詳しい要約
1. どんなもの?
2. 先行研究と比べてどこがすごい?
3. 技術・手法の肝は?
4. どうやって有効だと検証した?
5. 議論はある?
6. 次に読むべき論文は?
※ AIが要旨から生成した要約です。正確性は原文をご確認ください。
著者: Yujia Chen, Bowen Weng
分類: cs.RO
原文アブストラクト
Consider a typical test of a robot system: one observes a sequence of outcomes concerning some aspect of interest (crash or no crash, tracking error, time to completion), and reports a mean (crash risk, average error, mean time to completion) and, more importantly, an interval guaranteed to contain that mean at a prescribed confidence, referred to as a performance certificate. Given expensive real-world trials, the sample size is therefore small, and the certificate is often loose. Now consider the same procedure, except that before each real outcome is revealed, the operator ``peeks'' at a large bank of simulated results, and places a bet on where the real outcome will land. As the real outcomes settle the bets, the operator gains or loses wealth. One's ``trust'' over simulators also shifts within the portfolio. This paper develops that idea into a sim-to-real betting certificate framework with three contributions: (i) An algorithm that links a scalable bank of simulators to effective bets, and the accumulated betting wealth to the certificate. (ii) A proof that the returned certificate is anytime valid, covering the true mean with the prescribed probability, using any simulator bank. (iii) The guaranteed wealth-regret bounds yield configuration principles for the proposed algorithm and simulator bank design to deliver tight certificates. Experiments across synthetic distributions and real-world robot tests, covering both replayed standardized testing outcomes and online runtime evaluation, show the proposed method narrows the certificate by $51.6\%\pm16\%$ against classic and state-of-the-art baselines, and by $32.26\%\pm8\%$ in the extremely limited-sample regime ($\leq30$ samples).
関連論文
- 安全なシミュレーションから実世界への転移の証明可能な保証sim2real
- タスク関連特徴ダイナミクスの忠実度がロボット超音波走査のゼロショットsim-to-real転送を可能にするsim2real
- 実2シミュレーション動力学推定と強化学習によるトルク制御ロボットのSim2Real転送の強化sim2real
- 触覚シミュレーションを使わない触覚Sim2Real:ボトルネック潜在再構成による実現sim2real
- モデルミスマッチ下での計画の認証:乏しいデータからの到達可能性に関する三難問題sim2real
- LyEvO: リアプノフ誘導進化最適化による安全で堅牢なSim-to-Realポリシー学習sim2real