← all papers · overview

Can Unified Generation And Understanding Models Maintain Semantic Equivalence Across Different Output Modalities?

Abstract

Unified Multimodal Large Language Models (U-MLLMs) integrate understanding and generation within a single architecture. However, existing evaluations typically assess these capabilities separately, overlooking semantic equivalence, i.e., the ability to manifest consistent reasoning results regardless of the output modality. In this work, we investigate whether current U-MLLMs satisfy this premise.

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).