arXiv · 2609.39981
ExactVariance of Random Return in Distributional LQR and Its Application to Mean-Variance Optimal Control
Abstract
The classical linear quadratic regulator (LQR) minimizes the expected cumulative return but fails to account for performance variability, rendering it inadequate for risk-aware applications. To address this, we introduce the variance of the cumulative return as a risk measure in LQR. We derive the first exact closed-form expression for the variance of the discounted in?finite horizon return within the discrete-time Distributional LQR framework, for i.i.d. disturbances with symmetric probability densities. For Gaussian disturbances, this expression elegantly simplifies to a form dependent only on the disturbance covariance. Leveraging these theoretical foundations, we formulate a mean-variance optimal control problem that explicitly manages the trade-off between expected return and performance variability. To address the resulting non-convex optimization problem, we propose a novel adjoint gradient descent algorithm for a penalized formulation of the original problem, and establish that all iterates remain stabilizing and converge to a stationary point of the penalized objective. The effectiveness of this framework and the inherent risk-performance trade-off? are demonstrated through numerical experiments.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Ruyi Teng, Zifan Wang, Yulong Gao. 2026-09-30. ExactVariance of Random Return in Distributional LQR and Its Application to Mean-Variance Optimal Control. https://arxiv.org/abs/2609.39981
Cite the original work for its findings. Save a collection to share your selection of sources.