TY - RPRT TI - Learning When to Act: Communication-Efficient Reinforcement Learning via Run-Time Assurance AU - Adam Haroon AU - Erick J. Rodríguez-Seda AU - Cody Fleming AU - Tristan Schuler PY - 2026 UR - https://arxiv.org/abs/2605.12561 ID - 2605.12561 ER -