M-2KR
Emerging6papers using it
2024first seen
The M-2KR dataset/benchmark contains multimodal document collections with both images and text, and it is used to evaluate the performance of retrieval models on complex multimodal queries.
Papers using M-2KR (5)
- Recurrence-Enhanced Vision-and-Language Transformers for Robust
Multimodal Document RetrievalOverview of the EReL@MIR 2025 Multimodal Document Retrieval Challenge (Track 1)WikiSeeker: Rethinking the Role of Vision-Language Models in Knowledge-Based Visual Question AnsweringA Multi-Granularity Retrieval Framework for Visually-Rich DocumentsPreFLMR: Scaling Up Fine-Grained Late-Interaction Multi-modal Retrievers