TY - RPRT TI - Value Function Approximations via Kernel Embeddings for No-Regret Reinforcement Learning AU - Sayak Ray Chowdhury AU - Rafael Oliveira PY - 2022 UR - https://arxiv.org/abs/2011.07881 ID - 2011.07881 ER -