TY - RPRT TI - Constrained Reinforcement Learning with Average Reward Objective: Model-Based and Model-Free Algorithms AU - Vaneet Aggarwal AU - Washim Uddin Mondal AU - Qinbo Bai PY - 2024 UR - https://arxiv.org/abs/2406.11481 ID - 2406.11481 ER -