TY - RPRT TI - History Rhymes: Accelerating LLM Reinforcement Learning with RhymeRL AU - Jingkai He AU - Tianjian Li AU - Erhu Feng AU - Dong Du AU - Qian Liu AU - Tao Liu AU - Yubin Xia AU - Haibo Chen PY - 2025 UR - https://arxiv.org/abs/2508.18588 ID - 2508.18588 ER -