SEED-Bench
Canonical9papers using it
2023first seen
SEED-Bench Card Benchmark details Benchmark type: SEED-Bench is a large-scale benchmark to evaluate Multimodal Large Language Models (MLLMs). It consists of 19K multiple choice questions with accurate human annotations, which covers 12 evaluation dimensions including the comprehension of both the image and video modali
Papers using SEED-Bench (9)
- Make Your LVLM KV Cache More LightweightLLMind: Bio-inspired Training-free Adaptive Visual Representations for Vision-Language ModelsSame Answer, Different Representations: Hidden instability in VLMsVision to Geometry: 3D Spatial Memory for Sequential Embodied MLLM Reasoning and ExplorationHybridToken-VLM: Hybrid Token Compression for Vision-Language ModelsInternLM-XComposer: A Vision-Language Large Model for Advanced
Text-image Comprehension and CompositionEE-MLLM: A Data-Efficient and Compute-Efficient Multimodal Large
Language ModelDynamic Multimodal Evaluation with Flexible Complexity by Vision-Language BootstrappingExploring Multi-Grained Concept Annotations for Multimodal Large
Language Models