TY - RPRT TI - Beyond Single-Step Updates: Reinforcement Learning of Heuristics with Limited-Horizon Search AU - Gal Hadar AU - Forest Agostinelli AU - Shahaf S. Shperberg PY - 2025 UR - https://arxiv.org/abs/2511.10264 ID - 2511.10264 ER -