← all datasets

NoCaps

Canonical
4papers using it
2022first seen

Dubbed NoCaps, for novel object captioning at scale, NoCaps consists of 166,100 human-generated captions describing 15,100 images from the Open Images validation and test sets. The associated training data consists of COCO image-caption pairs, plus Open Images image-level labels and object bounding boxes. Since Open Im

Papers using NoCaps (4)

NoCaps dataset β€” papers, benchmarks & downloads Β· Multimodal