arXiv · 2511.13458
Trust in Vision-Language Models: Insights from a Participatory User Workshop
Abstract
With the growing deployment of Vision-Language Models (VLMs), pre-trained on large image-text and video-text datasets, it is critical to equip users with the tools to discern when to trust these systems. However, examining how user trust in VLMs builds and evolves remains an open problem. This problem is exacerbated by the increasing reliance on AI models as judges for experimental validation, to bypass the cost and implications of running participatory design studies directly with users. Following a user-centred approach, this paper presents preliminary results from a workshop with prospective VLM users. Insights from this pilot workshop inform future studies aimed at contextualising trust metrics and strategies for participants' engagement to fit the case of user-VLM interaction.
Explore related subjects
Keep this discovery
Agnese Chiatti, Lara Piccolo, Sara Bernardini, Matteo Matteucci, Viola Schiaffonati. 2025-11-17. Trust in Vision-Language Models: Insights from a Participatory User Workshop. https://arxiv.org/abs/2511.13458
Cite the original work for its findings. Save a collection to share your selection of sources.