← all papers · overview

Clean Evaluations On Contaminated Visual Language Models

Abstract

How to evaluate large language models (LLMs) cleanly has been established as an important research era to genuinely report the performance of possibly contaminated LLMs. Yet, how to cleanly evaluate the visual language models (VLMs) is an under-studied problem. We propose a novel approach to achieve such goals through data augmentation methods on the visual input information. We then craft a new v

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).