arXiv · 2312.08958
LiFT: Unsupervised Reinforcement Learning with Foundation Models as Teachers
Abstract
We propose a framework that leverages foundation models as teachers, guiding a reinforcement learning agent to acquire semantically meaningful behavior without human feedback. In our framework, the agent receives task instructions grounded in a training environment from large language models. Then, a vision-language model guides the agent in learning the multi-task language-conditioned policy by providing reward feedback. We demonstrate that our method can learn semantically meaningful skills in a challenging open-ended MineDojo environment while prior unsupervised skill discovery methods struggle. Additionally, we discuss observed challenges of using off-the-shelf foundation models as teachers and our efforts to address them.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Taewook Nam, Juyong Lee, Jesse Zhang, Sung Ju Hwang, Joseph J. Lim, Karl Pertsch. 2023-12-14. LiFT: Unsupervised Reinforcement Learning with Foundation Models as Teachers. https://arxiv.org/abs/2312.08958
Cite the original work for its findings. Save a collection to share your selection of sources.