ARC-AGI-3
Emerging7papers using it
2025first seen
The ARC-AGI-3 dataset/benchmark contains a set of game-playing challenges used to evaluate the performance and learning capabilities of large language model (LLM) agents in dynamic environments.
The ARC-AGI-3 dataset/benchmark contains a set of game-playing challenges used to evaluate the performance and learning capabilities of large language model (LLM) agents in dynamic environments.