TY - RPRT TI - Utilizing Maximum Mean Discrepancy Barycenter for Propagating the Uncertainty of Value Functions in Reinforcement Learning AU - Srinjoy Roy AU - Swagatam Das PY - 2024 UR - https://arxiv.org/abs/2404.00686 ID - 2404.00686 ER -