TY - RPRT TI - Full error analysis of policy gradient learning algorithms for exploratory linear quadratic mean-field control problem in continuous time with common noise AU - Noufel Frikha AU - Huyên Pham AU - Xuanye Song PY - 2024 UR - https://arxiv.org/abs/2408.02489 ID - 2408.02489 ER -