Memory-Based Reinforcement Learning: Efficient Computation with Prioritized Sweeping

Apr-6-2023, 19:07:18 GMT–Neural Information Processing Systems

We present a new algorithm, Prioritized Sweeping, for efficient prediction and control of stochastic Markov systems. Incremental learning methods such as Temporal Differencing and Q-Iearning have fast real time perfor(cid:173) mance. Classical methods are slower, but more accurate, because they make full use of the observations. Prioritized Sweeping aims for the best of both worlds. It uses all previous experiences both to prioritize impor(cid:173) tant dynamic programming sweeps and to guide the exploration of state(cid:173) space.

efficient computation, memory-based reinforcement learning, prioritized sweeping, (3 more...)

Neural Information Processing Systems

Apr-6-2023, 19:07:18 GMT

Conferences Web Page

Add feedback

Technology:
- Information Technology > Artificial Intelligence > Machine Learning > Reinforcement Learning (0.48)