TY - RPRT TI - Provably Efficient Reinforcement Learning with Multinomial Logit Function Approximation AU - Long-Fei Li AU - Yu-Jie Zhang AU - Peng Zhao AU - Zhi-Hua Zhou PY - 2025 UR - https://arxiv.org/abs/2405.17061 ID - 2405.17061 ER -