TY - RPRT TI - Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models AU - Ilias Kazantzidis AU - Timothy J. Norman AU - Yali Du AU - Christopher T. Freeman PY - 2026 UR - https://arxiv.org/abs/2607.13172 ID - 2607.13172 ER -