← all papers · overview

Genixer: Empowering Multimodal Large Language Models As A Powerful Data Generator

Abstract

Multimodal Large Language Models (MLLMs) demonstrate exceptional problem-solving capabilities, but few research studies aim to gauge the ability to generate visual instruction tuning data. This paper proposes to explore the potential of empowering MLLMs to generate data independently without relying on GPT-4. We introduce Genixer, a comprehensive data generation pipeline consisting of four key ste

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).