arXiv · 1703.07608
Deep Exploration via Randomized Value Functions
Abstract
We study the use of randomized value functions to guide deep exploration in reinforcement learning. This offers an elegant means for synthesizing statistically and computationally efficient exploration with common practical approaches to value function learning. We present several reinforcement learning algorithms that leverage randomized value functions and demonstrate their efficacy through computational studies. We also prove a regret bound that establishes statistical efficiency with a tabular representation.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Ian Osband, Benjamin Van Roy, Daniel Russo, Zheng Wen. 2019-09-23. Deep Exploration via Randomized Value Functions. https://arxiv.org/abs/1703.07608
Cite the original work for its findings. Save a collection to share your selection of sources.