arXiv · 2502.07978
A Survey of In-Context Reinforcement Learning
Abstract
Reinforcement learning (RL) agents typically optimize their policies by performing expensive backward passes to update their network parameters. However, some agents can solve new tasks without updating any parameters by simply conditioning on additional context such as their action-observation histories. This paper surveys work on such behavior, known as in-context reinforcement learning.
Explore related subjects
Keep this discovery
Amir Moeini, Jiuqi Wang, Jacob Beck, Ethan Blaser, Shimon Whiteson, Rohan Chandra, Shangtong Zhang. 2025-02-11. A Survey of In-Context Reinforcement Learning. https://arxiv.org/abs/2502.07978
Cite the original work for its findings. Save a collection to share your selection of sources.