arXiv · 1410.4604
Domain-Independent Optimistic Initialization for Reinforcement Learning
Abstract
In Reinforcement Learning (RL), it is common to use optimistic initialization of value functions to encourage exploration. However, such an approach generally depends on the domain, viz., the scale of the rewards must be known, and the feature representation must have a constant norm. We present a simple approach that performs optimistic initialization with less dependence on the domain.
Explore related subjects
Keep this discovery
Marlos C. Machado, Sriram Srinivasan, Michael Bowling. 2014-10-16. Domain-Independent Optimistic Initialization for Reinforcement Learning. https://arxiv.org/abs/1410.4604
Cite the original work for its findings. Save a collection to share your selection of sources.