TY - RPRT TI - Learning Near Optimal Policies with Low Inherent Bellman Error AU - Andrea Zanette AU - Alessandro Lazaric AU - Mykel Kochenderfer AU - Emma Brunskill PY - 2020 UR - https://arxiv.org/abs/2003.00153 ID - 2003.00153 ER -