TY - RPRT TI - Breaking the Performance Ceiling in Reinforcement Learning requires Inference Strategies AU - Felix Chalumeau AU - Daniel Rajaonarivonivelomanantsoa AU - Ruan de Kock AU - Claude Formanek AU - Sasha Abramowitz AU - Oumayma Mahjoub AU - Wiem Khlifi AU - Simon Du Toit AU - Louay Ben Nessir AU - Refiloe Shabe AU - Noah De Nicola AU - Arnol Fokam AU - Siddarth Singh AU - Ulrich Mbou Sob AU - Arnu Pretorius PY - 2025 UR - https://arxiv.org/abs/2505.21236 ID - 2505.21236 ER -