TY - RPRT TI - Revealing the Gap in Human and VLM Scene Perception through Counterfactual Semantic Saliency AU - Ziqi Wen AU - Parsa Madinei AU - Miguel P. Eckstein PY - 2026 UR - https://arxiv.org/abs/2605.13047 ID - 2605.13047 ER -