Antoine Cully
収録論文 33本 ・ フィジカルAI/ロボット学習
制御歩行進化計算進化計算/強化学習進化的方策探索強化学習
※arXiv著者名で収集。同姓同名の別人の論文が含まれる場合があります。
論文
- 船上クレーンの二重振り子揺れ抑制のためのオンボードMuJoCoベースモデル予測制御制御2026/3/1
船上クレーンの二重振り子揺れを抑制するため、MuJoCoシミュレータ上でクロスエントロピー法を用いたモデル予測制御をリアルタイムに実行するパイプラインを提案した。
- 遊びの時間:初期動物ダイナミクスのシミュレーションがロボットの移動発見を促進する歩行2025/9/1
動物の成長に倣い、ロボットのアクチュエータ強度を生涯にわたって変化させるカリキュラムSMOLを提案し、MAP-Elitesと組み合わせることで移動行動の性能と多様性を向上させる。
- 白紙状態から創発的能力へ:実世界の教師なし品質多様性によるロボットスキル発見歩行2025/8/1
実世界で四足歩行ロボットが多様な移動スキルを教師なしで自律的に発見・習得する手法URSAを提案し、損傷適応タスクでも有効性を示した。
- 教師なし品質多様性による欺瞞的な適応度最適化の克服進化計算2025/4/1
センサデータから特徴を学習する教師なし品質多様性アルゴリズムAURORAを改良し、ドメイン知識なしで欺瞞的な最適化問題を解けるようにした。
- 行動変異による大規模並列化で方策勾配品質多様性をスケーリング進化計算/強化学習2025/1/1
MAP-Elitesを方策勾配で拡張し、集中型actor-criticに依存せず大規模並列化可能にしたASCII-MEを提案。単一GPUで250秒未満に高性能な多様方策群を生成し、既存手法より平均5倍高速。
- ちょうどよい多様性で品質を高める進化的方策探索進化的方策探索2024/5/1
行動記述子と適合度の関係を学習し、有望な探索領域に評価を集中させることで、強化学習の進化的方策探索を効率化するJEDiフレームワークを提案。
- 品質多様性アクタークリティック:価値と後継特徴クリティックによる高性能で多様な行動の学習強化学習2024/3/15
価値関数と後継特徴の2つのクリティックを制約付き最適化で統合し、高いリターンと多様なスキルを同時に獲得するオフポリシー強化学習手法QDACを提案した。
- Synergizing Quality-Diversity with Descriptor-Conditioned Reinforcement Learning2024/1/1
- Mix-ME: Quality-Diversity for Multi-Agent Learning2023/11/3
- Quality-Diversity Optimisation on a Physical Robot Through Dynamics-Aware and Reset-Free Learning2023/4/1
- Don't Bet on Luck Alone: Enhancing Behavioral Reproducibility of Quality-Diversity Solutions in Uncertain Domains2023/4/1
- Understanding the Synergies between Quality-Diversity and Deep Reinforcement Learning2023/3/1
- Enhancing MAP-Elites with Multiple Parallel Evolution Strategies2023/3/1
- Improving the Data Efficiency of Multi-Objective Quality-Diversity through Gradient Assistance and Crowding Exploration2023/2/1
- Benchmarking Quality-Diversity Algorithms on Neuroevolution for Reinforcement Learning2022/11/1
- Discovering Unsupervised Behaviours from Full-State Trajectories2022/11/1
- Online Damage Recovery for Physical Robots with Hierarchical Quality-Diversity2022/10/1
- Efficient Learning of Locomotion Skills through the Discovery of Diverse Environmental Trajectory Generator Priors2022/10/1
- Relevance-guided Unsupervised Discovery of Abilities with Quality-Diversity Algorithms2022/4/1
- Hierarchical Quality-Diversity for Online Damage Recovery2022/4/1
- Learning to Walk Autonomously via Reset-Free Quality-Diversity2022/4/1
- Accelerated Quality-Diversity through Massive Parallelism2022/2/1
- Dynamics-Aware Quality-Diversity for Efficient Learning of Skill Repertoires2021/9/1
- Unsupervised Behaviour Discovery with Quality-Diversity Optimisation2021/6/1
- Fast and stable MAP-Elites in noisy domains using deep grids2020/6/1
- Multimodal representation models for prediction and control from partial information2019/10/1
- Autonomous skill discovery with Quality-Diversity and Unsupervised Descriptors2019/5/1
- Hierarchical Behavioral Repertoires with Unsupervised Descriptors2018/4/1
- Limbo: A Fast and Flexible Library for Bayesian Optimization2016/11/1
- Towards semi-episodic learning for robot damage recovery2016/10/1
- Robots that can adapt like animals2014/7/1
- Evolving a Behavioral Repertoire for a Walking Robot2013/8/1
- Fast Damage Recovery in Robotics with the T-Resilience Algorithm2013/2/1