arXiv · 2312.08958
LiFT: Unsupervised Reinforcement Learning with Foundation Models as Teachers
Abstract
We propose a framework that leverages foundation models as teachers, guiding a reinforcement learning agent to acquire semantically meaningful behavior without human feedback. In our framework, the agent receives task instructions grounded in a training environment from large language models. Then, a vision-language model guides the agent in learning the multi-task language-conditioned policy by providing reward feedback. We demonstrate that our method can learn semantically meaningful skills in a challenging open-ended MineDojo environment while prior unsupervised skill discovery methods struggle. Additionally, we discuss observed challenges of using off-the-shelf foundation models as teachers and our efforts to address them.
Explore related subjects
Keep this discovery
Taewook Nam, Juyong Lee, Jesse Zhang, Sung Ju Hwang, Joseph J. Lim, Karl Pertsch. 2023-12-14. LiFT: Unsupervised Reinforcement Learning with Foundation Models as Teachers. https://arxiv.org/abs/2312.08958
Cite the original work for its findings. Save a collection to share your selection of sources.