← all papers · overview

Sight Beyond Text: Multi-modal Training Enhances Llms In Truthfulness And Ethics

Abstract

Multi-modal large language models (MLLMs) are trained based on large language models (LLM), with an enhanced capability to comprehend multi-modal inputs and generate textual responses. While they excel in multi-modal tasks, the pure NLP abilities of MLLMs are often underestimated and left untested. In this study, we get out of the box and unveil an intriguing characteristic of MLLMs -- our prelimi

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).