SearcharxivSearch

arXiv subjects

Thibaut Bourdais

Publications and source records attributed to Thibaut Bourdais.

3 recordsLinked to original sources

Exponential twist of probability measures: drift correction in term of a generalized gradient

In this paper we study the exponential twist, i.e. a path-integral exponential change of measure, of a Markovian reference probability measure $¶$. This type of transformation naturally appears in variational representation formulae originating from the theory of large deviations and can be interpreted in some cases, as the solution of a specific stochastic control problem. Under a very general Markovian assumption on $¶$, we fully characterize the exponential twist probability measure as the solution of a martingale problem and prove that it inherits the Markov property of the reference measure. The ''generator'' of the martingale problem shows a drift depending on a {\it generalized gradient} of some suitable {\it value function} $v$. The analysis focuses on the fact that any Markovian probability measure fulfills an {\it intrinsic martingale problem} for which no uniqueness is required.

math.PR

An entropy penalized approach for stochastic control problems. Complete version

In this paper, we propose an original approach to stochastic control problems. We consider a weak formulation that is written as an optimization (minimization) problem on the space of probability measures. We then introduce a penalized version of this problem obtained by splitting the minimization variables and penalizing the discrepancy between the two variables via an entropy term. We show that the penalized problem provides a good approximation of the original problem when the weight of the entropy penalization term is large enough. Moreover, the penalized problem has the advantage of giving rise to two optimization subproblems that are easy to solve in each of the two optimization variables when the other is fixed. We take advantage of this property to propose an alternating optimization procedure that converges to the infimum of the penalized problem with a rate $O(1/k)$, where $k$ is the number of iterations. The relevance of this approach is illustrated by solving a high-dimensional stochastic control problem aimed at controlling consumption in electrical systems.

math.OC

An entropy penalized approach for stochastic optimization with marginal law constraints. Complete version

This paper focuses on stochastic optimal control problems with constraints in law, which are rewritten as optimization (minimization) of probability measures problem on the canonical space. We introduce a penalized version of this type of problems by splitting the optimization variable and adding an entropic penalization term. We prove that this penalized version constitutes a good approximation of the original control problem and we provide an alternating procedure which converges, under a so called ''Stability Condition'', to an approximate solution of the original problem. We extend the approach introduced in a previous paperof the same authors including a jump dynamics, non-convex costs and constraints on the marginal laws of the controlled process. The interest of our approach is illustrated by numerical simulations related to demand-side management problems arising in power systems.

math.OC