TY - RPRT TI - Value-Biased Maximum Likelihood Estimation for Model-based Reinforcement Learning in Discounted Linear MDPs AU - Yu-Heng Hung AU - Ping-Chun Hsieh AU - Akshay Mete AU - P. R. Kumar PY - 2023 UR - https://arxiv.org/abs/2310.11515 ID - 2310.11515 ER -