TY - RPRT TI - Enhancing Sample Efficiency and Exploration in Reinforcement Learning through the Integration of Diffusion Models and Proximal Policy Optimization AU - Tianci Gao AU - Konstantin A. Neusypin AU - Dmitry D. Dmitriev AU - Bo Yang AU - Shengren Rao PY - 2025 UR - https://arxiv.org/abs/2409.01427 ID - 2409.01427 ER -