MT-VQA
Emerging3papers using it
488HF downloads
42HF likes
2024first seen
Dataset Card The dataset is oriented toward visual question answering of multilingual text scenes in nine languages, including Korean, Japanese, Italian, Russian, Deutsch, French, Thai, Arabic, and Vietnamese. The question-answer pairs are labeled by native annotators following a series of rules. A comprehensive descri
π€ Hugging Faceβ cc-by-nc-4.0