TY - RPRT TI - Deployable Human Preference Alignment in Robotics: Learning Representative Rewards from Diverse Human Preferences AU - Taehyung Kim AU - Gwangmo Lee AU - Minjun Chang AU - Sunghyun Lim AU - Jongeun Choi PY - 2026 UR - https://arxiv.org/abs/2607.12466 ID - 2607.12466 ER -