TY - RPRT TI - REFINE-LM: Mitigating Language Model Stereotypes via Reinforcement Learning AU - Rameez Qureshi AU - Naïm Es-Sebbani AU - Luis Galárraga AU - Yvette Graham AU - Miguel Couceiro AU - Zied Bouraoui PY - 2024 UR - https://arxiv.org/abs/2408.09489 ID - 2408.09489 ER -