TY - RPRT TI - AI Alignment through Reinforcement Learning from Human Feedback? Contradictions and Limitations AU - Adam Dahlgren Lindström AU - Leila Methnani AU - Lea Krause AU - Petter Ericson AU - Íñigo Martínez de Rituerto de Troya AU - Dimitri Coelho Mollo AU - Roel Dobbe PY - 2024 UR - https://arxiv.org/abs/2406.18346 ID - 2406.18346 ER -