TY - RPRT TI - Linear Bellman Completeness Suffices for Efficient Online Reinforcement Learning with Few Actions AU - Noah Golowich AU - Ankur Moitra PY - 2024 UR - https://arxiv.org/abs/2406.11640 ID - 2406.11640 ER -