← all datasets

Video-Holmes

Emerging
3papers using it
68HF downloads
0HF likes
2025first seen

The 'Video-Holmes' dataset is a benchmark used to evaluate video reasoning tasks, specifically focusing on the performance of Multimodal Large Language Models (MLLMs) in generating reasoning clues through autonomous tool usage within chain-of-thought reasoning sequences.

Papers using Video-Holmes (3)

Video-Holmes dataset β€” papers, benchmarks & downloads Β· Multimodal