← all datasets

MT-VQA

Emerging
3papers using it
488HF downloads
42HF likes
2024first seen

Dataset Card The dataset is oriented toward visual question answering of multilingual text scenes in nine languages, including Korean, Japanese, Italian, Russian, Deutsch, French, Thai, Arabic, and Vietnamese. The question-answer pairs are labeled by native annotators following a series of rules. A comprehensive descri

Papers using MT-VQA (3)

MT-VQA dataset β€” papers, benchmarks & downloads Β· Multimodal