DAPO-Math-17K
Emerging3papers using it
12,330HF downloads
185HF likes
2025first seen
The 'DAPO-Math-17K' dataset/benchmark contains a collection of mathematical problems designed to evaluate the reasoning capabilities of Large Language Models (LLMs) in the context of Reinforcement Learning with Verifiable Rewards (RLVR).
π€ Hugging Faceβ apache-2.0