arXiv · 2402.18836
A Model-Based Approach for Improving Reinforcement Learning Efficiency Leveraging Expert Observations
Abstract
This paper investigates how to incorporate expert observations (without explicit information on expert actions) into a deep reinforcement learning setting to improve sample efficiency. First, we formulate an augmented policy loss combining a maximum entropy reinforcement learning objective with a behavioral cloning loss that leverages a forward dynamics model. Then, we propose an algorithm that automatically adjusts the weights of each component in the augmented loss function. Experiments on a variety of continuous control tasks demonstrate that the proposed algorithm outperforms various benchmarks by effectively utilizing available expert observations.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Erhan Can Ozcan, Vittorio Giammarino, James Queeney, Ioannis Ch. Paschalidis. 2024-02-29. A Model-Based Approach for Improving Reinforcement Learning Efficiency Leveraging Expert Observations. https://doi.org/10.1109/cdc56724.2024.10885921
Cite the original work for its findings. Save a collection to share your selection of sources.