← all datasets

DAPO-Math-17K

Emerging
3papers using it
12,330HF downloads
185HF likes
2025first seen

The 'DAPO-Math-17K' dataset/benchmark contains a collection of mathematical problems designed to evaluate the reasoning capabilities of Large Language Models (LLMs) in the context of Reinforcement Learning with Verifiable Rewards (RLVR).

Papers using DAPO-Math-17K (3)

DAPO-Math-17K dataset β€” papers, benchmarks & downloads Β· Reinforcement Learning