← all papers · overview

Finding And Editing Multi-modal Neurons In Pre-trained Transformers

Abstract

Understanding the internal mechanisms by which multi-modal large language models (LLMs) interpret different modalities and integrate cross-modal representations is becoming increasingly critical for continuous improvements in both academia and industry. In this paper, we propose a novel method to identify key neurons for interpretability -- how multi-modal LLMs bridge visual and textual concepts f

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).