TY - RPRT TI - Learning List-wise Representation in Reinforcement Learning for Ads Allocation with Multiple Auxiliary Tasks AU - Ze Wang AU - Guogang Liao AU - Xiaowen Shi AU - Xiaoxu Wu AU - Chuheng Zhang AU - Yongkang Wang AU - Xingxing Wang AU - Dong Wang PY - 2022 UR - https://arxiv.org/abs/2204.00888 ID - 2204.00888 ER -