TY - RPRT TI - Bellman Calibration for Marginalized Importance Weighting in Offline Reinforcement Learning AU - Lars van der Laan AU - Nathan Kallus PY - 2026 UR - https://arxiv.org/abs/2608.24858 ID - 2608.24858 ER -