arXiv · 1611.01224
Sample Efficient Actor-Critic with Experience Replay
Abstract
This paper presents an actor-critic deep reinforcement learning agent with experience replay that is stable, sample efficient, and performs remarkably well on challenging environments, including the discrete 57-game Atari domain and several continuous control problems. To achieve this, the paper introduces several innovations, including truncated importance sampling with bias correction, stochastic dueling network architectures, and a new trust region policy optimization method.
Explore related subjects
Keep this discovery
Ziyu Wang, Victor Bapst, Nicolas Heess, Volodymyr Mnih, Remi Munos, Koray Kavukcuoglu, Nando de Freitas. 2016-11-03. Sample Efficient Actor-Critic with Experience Replay. https://arxiv.org/abs/1611.01224
Cite the original work for its findings. Save a collection to share your selection of sources.