BIFROST: 観測空間のSim2Real転送のための不変特徴表現の橋渡し
BIFROST: Bridging Invariant Feature Representation for Observation-space Sim2Real Transfer
シミュレーションと実環境のギャップを埋めるため、クロスドメインの双模倣目的を用いて共有履歴エンコーダを学習し、ゼロショット転送を可能にする手法を提案した。
著者: Yunfu Deng, Josiah P. Hanna
分類: cs.RO, cs.LG
原文アブストラクト
Sim2real transfer for robot policy learning suffers due to mismatch between simulation and reality. Existing methods typically address each gap in isolation through separate adaptation modules, which are composed or layered when both gaps coexist. Yet the basis for attempting sim2real in the first place is that there is shared structure between a task in simulation and reality, where equivalent actions from equivalent configurations produce equivalent long term outcomes regardless of domain specific differences in rendering or physics. In this paper, we study whether we can identify and exploit this shared structure from raw observations to train a policy that enables zero shot transfer. We introduce BIFROST, which learns a shared history encoder on paired cross-domain data via cross-domain bisimulation objective: observation-action sequences leading to equivalent long-term behavior are mapped to nearby latent states, regardless of domain. Policies trained on these latent states in simulation transfer zero-shot to reality. We provide empirical evidence on sim2sim visual navigation and sim2real contact rich manipulation task and visual servoing task that BIFROST achieves effective transfer where domain adaptation and co-training baselines fail under both visual and dynamics domain gaps.
関連論文
- タスク関連特徴ダイナミクスの忠実度がロボット超音波走査のゼロショットsim-to-real転送を可能にするsim2real
- 実2シミュレーション動力学推定と強化学習によるトルク制御ロボットのSim2Real転送の強化sim2real
- シミュレーションから実世界への性能証明書のためのベッティングsim2real
- 触覚シミュレーションを使わない触覚Sim2Real:ボトルネック潜在再構成による実現sim2real
- R2S-EGO: スパースキャプチャ実世界からシミュレーションへのデュアルプロキシ精緻化sim2real
- LyEvO: リアプノフ誘導進化最適化による安全で堅牢なSim-to-Realポリシー学習sim2real