TY - RPRT TI - Distributional Reinforcement Learning with Dual Expectile-Quantile Regression AU - Sami Jullien AU - Romain Deffayet AU - Jean-Michel Renders AU - Paul Groth AU - Maarten de Rijke PY - 2025 UR - https://arxiv.org/abs/2305.16877 ID - 2305.16877 ER -