CyberSecEval
Emerging3papers using it
2023first seen
CyberSecEval The dataset source can be found here. (CyberSecEval2 Version) Abstract Large language models (LLMs) introduce new security risks, but there are few comprehensive evaluation suites to measure and reduce these risks. We present CYBERSECEVAL 2, a novel benchmark to quantify LLM security risks and capabilities