← all datasets

OK-VQA

Canonical
34papers using it
2021first seen

OK-VQA is a new dataset for visual question answering that requires methods which can draw upon outside knowledge to answer questions. - 14,055 open-ended questions - 5 ground truth answers per question - Manually filtered to ensure all questions require outside knowledge (e.g. from Wikipeida) - Reduced questions with

Papers using OK-VQA (34)

OK-VQA dataset — papers, benchmarks & downloads · Multimodal