TY - RPRT TI - Just-In-Time Reinforcement Learning: Continual Learning in LLM Agents Without Gradient Updates AU - Yibo Li AU - Zijie Lin AU - Ailin Deng AU - Xuan Zhang AU - Yufei He AU - Shuo Ji AU - Tri Cao AU - Bryan Hooi PY - 2026 UR - https://arxiv.org/abs/2601.18510 ID - 2601.18510 ER -