TY - RPRT TI - Clarify or Answer: Reinforcement Learning for Agentic VQA with Context Under-specification AU - Zongwan Cao AU - Bingbing Wen AU - Lucy Lu Wang PY - 2026 UR - https://arxiv.org/abs/2601.16400 ID - 2601.16400 ER -