← all datasets

CanItEdit

Emerging
3papers using it
754HF downloads
15HF likes
2025first seen

Can It Edit? Evaluating the Ability of Large Language Models to Follow Code Editing Instructions CanItEdit is a benchmark for evaluating LLMs on instructional code editing, the task of updating a program given a natural language instruction. The benchmark contains 105 hand-crafted Python programs with before and after

Papers using CanItEdit (3)

CanItEdit dataset β€” papers, benchmarks & downloads Β· AI for Code