Brandon Amos
収録論文 10本 ・ フィジカルAI/ロボット学習
強化学習/LLM報酬
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- 大規模言語モデルフィードバックによる意思決定エージェントのためのオンライン内的報酬強化学習/LLM報酬2024/10/1
LLMのフィードバックを非同期で利用し、RL方策と内的報酬関数を同時に学習する分散アーキテクチャONIを提案。NetHackで最先端性能を達成し、大規模オフラインデータセットを不要にした。
- Semi-Supervised Offline Reinforcement Learning with Action-Free Trajectories2022/10/1
- Theseus: A Library for Differentiable Nonlinear Optimization2022/7/1
- Nocturne: a scalable driving benchmark for bringing multi-agent learning one step closer to the real world2022/6/1
- Cross-Domain Imitation Learning via Optimal Transport2021/10/1
- On the model-based stochastic value gradient for continuous reinforcement learning2020/8/1
- Objective Mismatch in Model-based Reinforcement Learning2020/2/1
- Improving Sample Efficiency in Model-Free Reinforcement Learning from Images2019/10/1
- The Differentiable Cross-Entropy Method2019/9/1
- Learning Awareness Models2018/4/1