← all datasets

HumanEvalFix

Emerging
7papers using it
2023first seen

The 'HumanEvalFix' dataset/benchmark contains a collection of coding problems designed to evaluate the ability of models to generate unit tests that effectively reveal errors in faulty code while predicting correct outputs.

Papers using HumanEvalFix (7)

HumanEvalFix dataset β€” papers, benchmarks & downloads Β· AI for Code