TY - RPRT TI - Distilling Reinforcement Learning Policies for Interpretable Robot Locomotion: Gradient Boosting Machines and Symbolic Regression AU - Fernando Acero AU - Zhibin Li PY - 2024 UR - https://arxiv.org/abs/2403.14328 ID - 2403.14328 ER -