← all datasets

BUILD-BENCH

Emerging
3papers using it
2025first seen

Build-bench is an end-to-end benchmark that evaluates the capability of large language models to repair build failures in cross-instruction set architecture (ISA) settings using a collection of 268 real-world failed packages and auxiliary tools for autonomous reasoning.

Papers using BUILD-BENCH (3)

BUILD-BENCH dataset β€” papers, benchmarks & downloads Β· AI for Code