TY - RPRT TI - Rollout Efficiency in Reinforcement Learning for Reasoning Large Language Models: A Taxonomy and Future Directions AU - Niloofar Gholipour AU - Marcos Assuncao AU - Gursimran Singh AU - Timothy Yu AU - Rajkumar Buyya AU - Julien Gascon-Samson AU - Zhenan Fan AU - Yong Zhang AU - Xiaojie Xu AU - Yaqiang Yao AU - Xiaolong Bai PY - 2026 UR - https://arxiv.org/abs/2609.25463 ID - 2609.25463 ER -