TY - RPRT TI - PREDILECT: Preferences Delineated with Zero-Shot Language-based Reasoning in Reinforcement Learning AU - Simon Holk AU - Daniel Marta AU - Iolanda Leite PY - 2024 DO - 10.1145/3610977.3634970 UR - https://arxiv.org/abs/2402.15420 ID - 2402.15420 ER -