TY - RPRT TI - Can Large Vision-Language Models Correct Semantic Grounding Errors By Themselves? AU - Yuan-Hong Liao AU - Rafid Mahmood AU - Sanja Fidler AU - David Acuna PY - 2025 UR - https://arxiv.org/abs/2404.06510 ID - 2404.06510 ER -