TY - RPRT TI - Self-correcting Reward Shaping via Language Models for Reinforcement Learning Agents in Games AU - António Afonso AU - Iolanda Leite AU - Alessandro Sestini AU - Florian Fuchs AU - Konrad Tollmar AU - Linus Gisslén PY - 2025 UR - https://arxiv.org/abs/2506.23626 ID - 2506.23626 ER -