TY - RPRT TI - IPD: Boosting Sequential Policy with Imaginary Planning Distillation in Offline Reinforcement Learning AU - Yihao Qin AU - Yuanfei Wang AU - Hang Zhou AU - Peiran Liu AU - Hao Dong AU - Yiding Ji PY - 2026 UR - https://arxiv.org/abs/2603.04289 ID - 2603.04289 ER -