Memory-Based Reinforcement Learning: Efficient Computation with Prioritized Sweeping

Moore, Andrew W., Atkeson, Christopher G.

Dec-31-1993–Neural Information Processing Systems

We present a new algorithm, Prioritized Sweeping, for efficient prediction and control of stochastic Markov systems. Incremental learning methods such as Temporal Differencing and Q-Iearning have fast real time performance. Classical methods are slower, but more accurate, because they make full use of the observations. Prioritized Sweeping aims for the best of both worlds. It uses all previous experiences both to prioritize important dynamic programming sweeps and to guide the exploration of statespace. We compare Prioritized Sweeping with other reinforcement learning schemes for a number of different stochastic optimal control problems. It successfully solves large state-space real time problems with which other methods have difficulty.

learning, prioritized sweeping, reinforcement learning, (12 more...)

Neural Information Processing Systems

Dec-31-1993

Conferences PDF

Add feedback

Country:
- North America > United States
  - Massachusetts > Middlesex County > Cambridge (0.15)
- Europe > United Kingdom
  - England > Cambridgeshire > Cambridge (0.04)

Technology:
- Information Technology > Artificial Intelligence
  - Representation & Reasoning (1.00)
  - Machine Learning > Reinforcement Learning (1.00)

Duplicate Docs Excel Report

Title
Memory-Based Reinforcement Learning: Efficient Computation with Prioritized Sweeping
Memory-Based Reinforcement Learning: Efficient Computation with Prioritized Sweeping

Similar Docs Excel Report more

Title	Similarity	Source
None found