AdvBench
Emerging2papers using it
2025first seen
AdvBench is a benchmark used to evaluate the safety and trustworthiness of large language models by assessing their responses to potentially harmful content.
AdvBench is a benchmark used to evaluate the safety and trustworthiness of large language models by assessing their responses to potentially harmful content.