TY - RPRT TI - Re-FORC: Adaptive Reward Prediction for Efficient Chain-of-Thought Reasoning AU - Renos Zabounidis AU - Aditya Golatkar AU - Michael Kleinman AU - Alessandro Achille AU - Wei Xia AU - Stefano Soatto PY - 2026 UR - https://arxiv.org/abs/2511.02130 ID - 2511.02130 ER -