Flickr30k
Emerging3papers using it
2019first seen
The 'Flickr 30K' dataset contains 30,000 images, each paired with five descriptive captions, and is used to evaluate image captioning models by assessing their ability to generate coherent and semantically meaningful textual descriptions for given images.