TY - RPRT TI - Relative Importance Sampling for off-Policy Actor-Critic in Deep Reinforcement Learning AU - Mahammad Humayoo AU - Gengzhong Zheng AU - Xiaoqing Dong AU - Liming Miao AU - Shuwei Qiu AU - Zexun Zhou AU - Peitao Wang AU - Zakir Ullah AU - Naveed Ur Rehman Junejo AU - Xueqi Cheng PY - 2025 DO - 10.1038/s41598-025-96201-5 UR - https://arxiv.org/abs/1810.12558 ID - 1810.12558 ER -