← all datasets

MME

Canonical
33papers using it
2023first seen

The 'MME' dataset/benchmark is used to evaluate the performance of Multi-modal Large Language Models (MLLMs) in mitigating visual hallucination by comparing their generated responses against accurate visual cues from provided images.

Papers using MME (33)

MME dataset β€” papers, benchmarks & downloads Β· Multimodal