← all papers · overview

Mapping User Trust In Vision Language Models: Research Landscape, Challenges, And Prospects

Abstract

The rapid adoption of Vision Language Models (VLMs), pre-trained on large image-text and video-text datasets, calls for protecting and informing users about when to trust these systems. This survey reviews studies on trust dynamics in user-VLM interactions, through a multi-disciplinary taxonomy encompassing different cognitive science capabilities, collaboration modes, and agent behaviours. Literature insights and findings from a workshop with prospective VLM users inform preliminary requirements for future VLM trust studies.

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).