← all datasets

LiveCodeBench

Emerging
17papers using it
2025first seen

The 'LiveCodeBench' dataset/benchmark contains a collection of coding tasks and is used to evaluate the performance of large language models in generating correct and robust code solutions.

Papers using LiveCodeBench (17)

LiveCodeBench dataset β€” papers, benchmarks & downloads Β· AI Agents