TY - RPRT TI - Stratifying Reinforcement Learning with Signal Temporal Logic AU - Justin Curry AU - Alberto Speranzon PY - 2026 UR - https://arxiv.org/abs/2604.04923 ID - 2604.04923 ER -