TY - RPRT TI - Policy gradient learning methods for stochastic control with exit time and applications to share repurchase pricing AU - Mohamed Hamdouche AU - Pierre Henry-Labordere AU - Huyen Pham PY - 2023 UR - https://arxiv.org/abs/2302.07320 ID - 2302.07320 ER -