arXiv · 2507.12657
Distributional Reinforcement Learning on Path-dependent Options
Abstract
We reinterpret and propose a framework for pricing path-dependent financial derivatives by estimating the full distribution of payoffs using Distributional Reinforcement Learning (DistRL). Unlike traditional methods that focus on expected option value, our approach models the entire conditional distribution of payoffs, allowing for risk-aware pricing, tail-risk estimation, and enhanced uncertainty quantification. We demonstrate the efficacy of this method on Asian options, using quantile-based value function approximators.
Explore related subjects
Keep this discovery
Ahmet Umur Özsoy. 2025-07-16. Distributional Reinforcement Learning on Path-dependent Options. https://arxiv.org/abs/2507.12657
Cite the original work for its findings. Save a collection to share your selection of sources.