← all datasets

SEED-Bench

Canonical
9papers using it
2023first seen

SEED-Bench Card Benchmark details Benchmark type: SEED-Bench is a large-scale benchmark to evaluate Multimodal Large Language Models (MLLMs). It consists of 19K multiple choice questions with accurate human annotations, which covers 12 evaluation dimensions including the comprehension of both the image and video modali

Papers using SEED-Bench (9)

SEED-Bench dataset β€” papers, benchmarks & downloads Β· Multimodal