Do Agents Dream of Electric Sheep?: Improving Generalization in Reinforcement Learning through Generative Learning
Franceschelli, Giorgio, Musolesi, Mirco
–arXiv.org Artificial Intelligence
The Overfitted Brain hypothesis (Hoel, 2021) suggests dreams happen to allow generalization in the human brain. Here, we ask if the same is true for reinforcement learning agents as well. Given limited experience in a real environment, we use imagination-based reinforcement learning to train a policy on dream-like episodes, where non-imaginative, predicted trajectories are modified through generative augmentations. Experiments on four ProcGen environments show that, compared to classic imagination and offline training on collected experience, our method can reach a higher level of generalization when dealing with sparsely rewarded environments.
arXiv.org Artificial Intelligence
Mar-12-2024
- Country:
- Europe
- United Kingdom > England
- Greater London > London (0.04)
- Italy > Emilia-Romagna
- Metropolitan City of Bologna > Bologna (0.04)
- United Kingdom > England
- Asia > Middle East
- Jordan (0.04)
- Europe
- Genre:
- Research Report (0.40)
- Technology: