TY - RPRT TI - A Note on Hybrid Online Reinforcement and Imitation Learning for LLMs: Formulations and Algorithms AU - Yingru Li AU - Ziniu Li AU - Jiacai Liu PY - 2025 UR - https://arxiv.org/abs/2512.23097 ID - 2512.23097 ER -