TY - RPRT TI - One-Step Generative Policies with Q-Learning: A Reformulation of MeanFlow AU - Zeyuan Wang AU - Da Li AU - Yulin Chen AU - Ye Shi AU - Liang Bai AU - Tianyuan Yu AU - Yanwei Fu PY - 2025 UR - https://arxiv.org/abs/2511.13035 ID - 2511.13035 ER -