TY - RPRT TI - Self-Play Reinforcement Learning under Imperfect Information in Big 2 AU - Aalok Patwa PY - 2026 UR - https://arxiv.org/abs/2605.28863 ID - 2605.28863 ER -