TY - RPRT TI - Foundations of Safe Online Reinforcement Learning in the Linear Quadratic Regulator: $\sqrt{T}$-Regret AU - Benjamin Schiffer AU - Lucas Janson PY - 2025 UR - https://arxiv.org/abs/2504.18657 ID - 2504.18657 ER -