TY - RPRT TI - Reinforcement Learning for Optimal Stopping in POMDPs with Application to Quickest Change Detection AU - Austin Cooper AU - Sean Meyn PY - 2025 UR - https://arxiv.org/abs/2512.22347 ID - 2512.22347 ER -