TY - RPRT TI - Issues with Value-Based Multi-objective Reinforcement Learning: Value Function Interference and Overestimation Sensitivity AU - Peter Vamplew AU - Ethan AU - Watkins AU - Cameron Foale AU - Richard Dazeley PY - 2026 UR - https://arxiv.org/abs/2402.06266 ID - 2402.06266 ER -