TY - RPRT TI - Reinforcement Learning from LLM Feedback to Counteract Goal Misgeneralization AU - Houda Nait El Barj AU - Theophile Sautory PY - 2024 UR - https://arxiv.org/abs/2401.07181 ID - 2401.07181 ER -