Replication of Multi-agent Reinforcement Learning for the "Hide and Seek" Problem

Kamal, Haider, Niazi, Muaz A., Afzal, Hammad

Oct-9-2023–arXiv.org Artificial Intelligence

Reinforcement learning generates policies based on reward functions and hyperparameters. Slight changes in these can significantly affect results. The lack of documentation and reproducibility in Reinforcement learning research makes it difficult to replicate once-deduced strategies. While previous research has identified strategies using grounded maneuvers, there is limited work in more complex environments. The agents in this study are simulated similarly to Open Al's hider and seek agents, in addition to a flying mechanism, enhancing their mobility, and expanding their range of possible actions and strategies. This added functionality improves the Hider agents to develop a chasing strategy from approximately 2 million steps to 1.6 million steps and hiders

agent, multi-agent reinforcement learning, reinforcement learning, (8 more...)

arXiv.org Artificial Intelligence

Oct-9-2023

arXiv.org PDF

Add feedback

Country:
- North America
  - United States > Massachusetts
    - Middlesex County > Cambridge (0.04)
  - Canada > British Columbia
    - Metro Vancouver Regional District > Vancouver (0.04)
- Europe
  - Switzerland > Basel-City
    - Basel (0.04)
  - Latvia > Riga Municipality
    - Riga (0.04)
- Asia > Pakistan
  - Islamabad Capital Territory > Islamabad (0.04)

Genre:
- Research Report > New Finding (0.34)

Industry:
- Leisure & Entertainment > Games > Computer Games (0.93)

Technology:
- Information Technology > Artificial Intelligence
  - Representation & Reasoning > Agents (1.00)
  - Robots > Autonomous Vehicles
    - Drones (0.46)
  - Machine Learning
    - Reinforcement Learning (1.00)
    - Neural Networks > Deep Learning (0.46)

Duplicate Docs Excel Report

Title
None found

Similar Docs Excel Report more

Title	Similarity	Source
None found