ClawBench
Emerging4papers using it
2026first seen
'ClawBench' is a benchmark that contains a set of execution tasks used to evaluate the performance of large language model (LLM) agent systems in terms of their routing and execution capabilities.
'ClawBench' is a benchmark that contains a set of execution tasks used to evaluate the performance of large language model (LLM) agent systems in terms of their routing and execution capabilities.