← all datasets

ClawBench

Emerging
4papers using it
2026first seen

'ClawBench' is a benchmark that contains a set of execution tasks used to evaluate the performance of large language model (LLM) agent systems in terms of their routing and execution capabilities.

Papers using ClawBench (4)

ClawBench dataset β€” papers, benchmarks & downloads Β· AI Agents