Video-Holmes
Emerging3papers using it
68HF downloads
0HF likes
2025first seen
The 'Video-Holmes' dataset is a benchmark used to evaluate video reasoning tasks, specifically focusing on the performance of Multimodal Large Language Models (MLLMs) in generating reasoning clues through autonomous tool usage within chain-of-thought reasoning sequences.
π€ Hugging Faceβ apache-2.0