← all papers · overview

Sycophantasy: Quantifying Sycophancy And Hallucination In Small Open Weight Vlms For Vision-language Scoring Of Fantasy Characters

Abstract

Vision-language models (VLMs) are increasingly deployed as evaluators in tasks requiring nuanced image understanding, yet their reliability in scoring alignment between images and text descriptions remains underexplored. We investigate whether small, open-weight VLMs exhibit *sycophantic* behavior when evaluating image-text alignment: assigning high scores without grounding their judgments in visual evidence. To quantify this phenomenon, we introduce the *Bluffing Coefficient* (\bc), a metric that measures the mismatch between a model's score and its evidence recall. We evaluate six open-weight VLMs ranging from 450M to 8B parameters on a benchmark of 173,810 AI-generated character portraits paired with detailed textual descriptions. Our analysis reveals a significant inverse correlation between model size and sycophancy rate (, ), with smaller models exhibiting substantially higher rates of unjustified high s

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).