TY - RPRT TI - RRM: Robust Reward Model Training Mitigates Reward Hacking AU - Tianqi Liu AU - Wei Xiong AU - Jie Ren AU - Lichang Chen AU - Junru Wu AU - Rishabh Joshi AU - Yang Gao AU - Jiaming Shen AU - Zhen Qin AU - Tianhe Yu AU - Daniel Sohn AU - Anastasiia Makarova AU - Jeremiah Liu AU - Yuan Liu AU - Bilal Piot AU - Abe Ittycheriah AU - Aviral Kumar AU - Mohammad Saleh PY - 2025 UR - https://arxiv.org/abs/2409.13156 ID - 2409.13156 ER -