← all papers · overview

Fact :teaching Mllms With Faithful, Concise And Transferable Rationales

Abstract

The remarkable performance of Multimodal Large Language Models (MLLMs) has unequivocally demonstrated their proficient understanding capabilities in handling a wide array of visual tasks. Nevertheless, the opaque nature of their black-box reasoning processes persists as an enigma, rendering them uninterpretable and struggling with hallucination. Their ability to execute intricate compositional rea

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).