TY - RPRT TI - Position: Deployed Reinforcement Learning should be Continual AU - Parnian Behdin AU - Kevin Roice AU - Golnaz Mesbahi PY - 2026 UR - https://arxiv.org/abs/2606.04029 ID - 2606.04029 ER -