TY - RPRT TI - Evaluating Interpretable Reinforcement Learning by Distilling Policies into Programs AU - Hector Kohler AU - Quentin Delfosse AU - Waris Radji AU - Riad Akrour AU - Philippe Preux PY - 2025 UR - https://arxiv.org/abs/2503.08322 ID - 2503.08322 ER -