TY - RPRT TI - Reinforcement Learning via Parametric Cost Function Approximation for Multistage Stochastic Programming AU - Saeed Ghadimi AU - Raymond T. Perkins AU - Warren B. Powell PY - 2020 UR - https://arxiv.org/abs/2001.00831 ID - 2001.00831 ER -