Domain-Independent Optimistic Initialization for Reinforcement Learning

Machado, Marlos C. (University of Alberta) | Srinivasan, Sriram (University of Alberta) | Bowling, Michael (University of Alberta)

AAAI Conferences 

In Reinforcement Learning, it is common to use optimistic initialization of value functions to encourage exploration. However, such an approach generally depends on the domain, viz., the scale of the rewards must be known, and, when using function approximation, the feature representation must have a constant norm. We present a simple approach that performs optimistic initialization with less dependence on the domain.

Duplicate Docs Excel Report

Title
None found

Similar Docs  Excel Report  more

TitleSimilaritySource
None found