TY - RPRT TI - Sparse Mixture-of-Experts Reward Models Learn Interpretable and Specialized Experts for Personalized Preference Modeling AU - Yifan Wang AU - Jinyi Mu AU - Mayank Jobanputra AU - Yu Wang AU - Ji-Ung Lee AU - Soyoung Oh AU - Isabel Valera AU - Vera Demberg PY - 2026 UR - https://arxiv.org/abs/2606.04284 ID - 2606.04284 ER -