TY - RPRT TI - Utility-Based Reinforcement Learning: Unifying Single-objective and Multi-objective Reinforcement Learning AU - Peter Vamplew AU - Cameron Foale AU - Conor F. Hayes AU - Patrick Mannion AU - Enda Howley AU - Richard Dazeley AU - Scott Johnson AU - Johan Källström AU - Gabriel Ramos AU - Roxana Rădulescu AU - Willem Röpke AU - Diederik M. Roijers PY - 2024 UR - https://arxiv.org/abs/2402.02665 ID - 2402.02665 ER -