← all datasets

AetherCode

Emerging
3papers using it
2025first seen

AetherCode: Evaluating LLMs' Ability to Win In Premier Programming Competitions Introduction Competitive programming has emerged as a critical benchmark for evaluating the reasoning and coding capabilities of Large Language Models (LLMs). Despite impressive progress on existing benchmarks, we argue that current evaluat

Papers using AetherCode (3)

AetherCode dataset β€” papers, benchmarks & downloads Β· AI for Code