TY - RPRT TI - Flow-based Policy With Distributional Reinforcement Learning in Trajectory Optimization AU - Ruijie Hao AU - Longfei Zhang AU - Yang Dai AU - Yang Ma AU - Xingxing Liang AU - Guangquan Cheng PY - 2026 UR - https://arxiv.org/abs/2604.00977 ID - 2604.00977 ER -