← all datasets

VCR

Emerging
12papers using it
2019first seen

The VCR (Visual Commonsense Reasoning) dataset is used to evaluate a model's reasoning ability in understanding the semantics of visual content and natural language through tasks that require fine-grained visual and textual information.

Papers using VCR (12)

VCR dataset β€” papers, benchmarks & downloads Β· Multimodal