TY - RPRT TI - Stable and Safe Human-aligned Reinforcement Learning through Neural Ordinary Differential Equations AU - Liqun Zhao AU - Keyan Miao AU - Konstantinos Gatsis AU - Antonis Papachristodoulou PY - 2024 UR - https://arxiv.org/abs/2401.13148 ID - 2401.13148 ER -