← all papers · overview

A Concept-based Explainability Framework For Large Multimodal Models

Abstract

Large multimodal models (LMMs) combine unimodal encoders and large language models (LLMs) to perform multimodal tasks. Despite recent advancements towards the interpretability of these models, understanding internal representations of LMMs remains largely a mystery. In this paper, we present a novel framework for the interpretation of LMMs. We propose a dictionary learning based approach, applied

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).