TY - RPRT TI - A Policy Efficient Reduction Approach to Convex Constrained Deep Reinforcement Learning AU - Tianchi Cai AU - Wenpeng Zhang AU - Lihong Gu AU - Xiaodong Zeng AU - Jinjie Gu PY - 2022 UR - https://arxiv.org/abs/2108.12916 ID - 2108.12916 ER -