TY - RPRT TI - Efficient Learning of POMDPs with Known Observation Model in Average-Reward Setting AU - Alessio Russo AU - Alberto Maria Metelli AU - Marcello Restelli PY - 2024 UR - https://arxiv.org/abs/2410.01331 ID - 2410.01331 ER -