TY - RPRT TI - Understanding the theoretical properties of projected Bellman equation, linear Q-learning, and approximate value iteration AU - Han-Dong Lim AU - Donghwan Lee PY - 2025 UR - https://arxiv.org/abs/2504.10865 ID - 2504.10865 ER -