TY - RPRT TI - NADPEx: An on-policy temporally consistent exploration method for deep reinforcement learning AU - Sirui Xie AU - Junning Huang AU - Lanxin Lei AU - Chunxiao Liu AU - Zheng Ma AU - Wei Zhang AU - Liang Lin PY - 2018 UR - https://arxiv.org/abs/1812.09028 ID - 1812.09028 ER -