TY - RPRT TI - On the Sample Complexity of Discounted Reinforcement Learning with Optimized Certainty Equivalents AU - Oliver Mortensen AU - Mohammad Sadegh Talebi PY - 2026 UR - https://arxiv.org/abs/2605.21763 ID - 2605.21763 ER -