← all papers · overview

VL-ICL Bench: The Devil In The Details Of Multimodal In-context Learning

Abstract

Large language models (LLMs) famously exhibit emergent in-context learning (ICL) -- the ability to rapidly adapt to new tasks using few-shot examples provided as a prompt, without updating the model's weights. Built on top of LLMs, vision large language models (VLLMs) have advanced significantly in areas such as recognition, reasoning, and grounding. However, investigations into *multimodal ICL* h

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).