TY - RPRT TI - FAST-Q: Fast-track Exploration with Adversarially Balanced State Representations for Counterfactual Action Estimation in Offline Reinforcement Learning AU - Pulkit Agrawal AU - Rukma Talwadker AU - Aditya Pareek AU - Tridib Mukherjee PY - 2025 DO - 10.1145/3701716.3715224 UR - https://arxiv.org/abs/2504.21383 ID - 2504.21383 ER -