TY - RPRT TI - Provable and Practical: Efficient Exploration in Reinforcement Learning via Langevin Monte Carlo AU - Haque Ishfaq AU - Qingfeng Lan AU - Pan Xu AU - A. Rupam Mahmood AU - Doina Precup AU - Anima Anandkumar AU - Kamyar Azizzadenesheli PY - 2024 UR - https://arxiv.org/abs/2305.18246 ID - 2305.18246 ER -