日本フィジカルAI新聞

世界のフィジカルAIを、日本語で。

週刊ニュースレター購読
arXiv:2310.05430

Replication of Multi-agent Reinforcement Learning for the "Hide and Seek" Problem

Replication of Multi-agent Reinforcement Learning for the "Hide and Seek" Problem

シェア:XThreadsFacebookLINEはてブBluesky

著者: Haider Kamal, Muaz A. Niazi, Hammad Afzal

分類: cs.AI, cs.LG, cs.MA, cs.RO

原文アブストラクト

Reinforcement learning generates policies based on reward functions and hyperparameters. Slight changes in these can significantly affect results. The lack of documentation and reproducibility in Reinforcement learning research makes it difficult to replicate once-deduced strategies. While previous research has identified strategies using grounded maneuvers, there is limited work in more complex environments. The agents in this study are simulated similarly to Open Al's hider and seek agents, in addition to a flying mechanism, enhancing their mobility, and expanding their range of possible actions and strategies. This added functionality improves the Hider agents to develop a chasing strategy from approximately 2 million steps to 1.6 million steps and hiders