TY - RPRT TI - Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts AU - Yueqin Yin AU - Zhendong Wang AU - Yi Gu AU - Hai Huang AU - Weizhu Chen AU - Mingyuan Zhou PY - 2024 UR - https://arxiv.org/abs/2402.10958 ID - 2402.10958 ER -