Claude Sonnet 4.5
Emerging1papers using it
2025first seen
The 'Claude-Sonnet-4.5' dataset/benchmark is used to evaluate the performance of coding agents in the context of tool invocation and security vulnerabilities.
The 'Claude-Sonnet-4.5' dataset/benchmark is used to evaluate the performance of coding agents in the context of tool invocation and security vulnerabilities.