TY - RPRT TI - Mitigating Hallucinations in Large Vision-Language Models via DPO: On-Policy Data Hold the Key AU - Zhihe Yang AU - Xufang Luo AU - Dongqi Han AU - Yunjian Xu AU - Dongsheng Li PY - 2025 UR - https://arxiv.org/abs/2501.09695 ID - 2501.09695 ER -