← all papers · overview

Quantifying And Mitigating Unimodal Biases In Multimodal Large Language Models: A Causal Perspective

Abstract

Recent advancements in Large Language Models (LLMs) have facilitated the development of Multimodal LLMs (MLLMs). Despite their impressive capabilities, MLLMs often suffer from over-reliance on unimodal biases (e.g., language bias and vision bias), leading to incorrect answers or hallucinations in complex multimodal tasks. To investigate this issue, we propose a causal framework to interpret the bi

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).