TY - RPRT TI - Risk-Sensitive Reinforcement Learning with Smoothed Quantile Objectives AU - Mohammad Alipour-Vaezi AU - Huaiyang Zhong AU - Sajad Khodadadian PY - 2026 UR - https://arxiv.org/abs/2608.22227 ID - 2608.22227 ER -