Convergent Combinations of Reinforcement Learning with Linear Function Approximation

Dec-31-2003–Neural Information Processing Systems

Convergence for iterative reinforcement learning algorithms like TD(O) depends on the sampling strategy for the transitions. However, in practical applications it is convenient to take transition data from arbitrary sources without losing convergence. In this paper we investigate the problem of repeated synchronous updates based on a fixed set of transitions. Our main theorem yields sufficient conditions of convergence for combinations of reinforcement learning algorithms and linear function approximation. This allows to analyse if a certain reinforcement learning algorithm and a certain function approximator are compatible.

algorithm, artificial intelligence, fuzzy logic, (16 more...)

Neural Information Processing Systems

Dec-31-2003

Conferences PDF

Add feedback

Country:
- Europe > Germany (0.28)
- North America > United States
  - Massachusetts (0.14)

Technology:
- Information Technology > Artificial Intelligence
  - Machine Learning > Reinforcement Learning (1.00)
  - Representation & Reasoning > Uncertainty
    - Fuzzy Logic (0.64)

Duplicate Docs Excel Report

Title
Convergent Combinations of Reinforcement Learning with Linear Function Approximation
Convergent Combinations of Reinforcement Learning with Linear Function Approximation

Similar Docs Excel Report more

Title	Similarity	Source
None found