TY - RPRT TI - Infinite-Horizon Reinforcement Learning with Multinomial Logistic Function Approximation AU - Jaehyun Park AU - Junyeop Kwon AU - Dabeen Lee PY - 2024 UR - https://arxiv.org/abs/2406.13633 ID - 2406.13633 ER -