TY - RPRT TI - Contextual inference from single objects in Vision-Language models AU - Martina G. Vilas AU - Timothy Schaumlöffel AU - Gemma Roig PY - 2026 UR - https://arxiv.org/abs/2603.26731 ID - 2603.26731 ER -