Learning to Play in a Day: Faster Deep Reinforcement Learning by Optimality Tightening

He, Frank S., Liu, Yang, Schwing, Alexander G., Peng, Jian

Nov-5-2016–arXiv.org Machine Learning

We propose a novel training algorithm for reinforcement learning which combines the strength of deep Q-learning with a constrained optimization approach to tighten optimality and encourage faster reward propagation. Our novel technique makes deep reinforcement learning more practical by drastically reducing the training time. We evaluate the performance of our approach on the 49 games of the challenging Arcade Learning Environment, and report significant improvements in both training time and accuracy.

artificial intelligence, machine learning, reinforcement learning, (14 more...)

arXiv.org Machine Learning

Nov-5-2016

arXiv.org PDF

Add feedback

Country:
- Europe > United Kingdom > England (0.28)

Genre:
- Research Report (0.85)

Industry:
- Leisure & Entertainment > Games (1.00)

Technology:
- Information Technology > Artificial Intelligence > Machine Learning
  - Reinforcement Learning (1.00)
  - Neural Networks > Deep Learning (0.47)

Duplicate Docs Excel Report

Title
None found

Similar Docs Excel Report more

Title	Similarity	Source
None found