← all papers · overview

Understanding Alignment In Multimodal Llms: A Comprehensive Study

Abstract

Preference alignment has become a crucial component in enhancing the performance of Large Language Models (LLMs), yet its impact in Multimodal Large Language Models (MLLMs) remains comparatively underexplored. Similar to language models, MLLMs for image understanding tasks encounter challenges like hallucination. In MLLMs, hallucination can occur not only by stating incorrect facts but also by pro

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).