CoSApien
Emerging2papers using it
41HF downloads
1HF likes
2024first seen
CoSApien: A Human-Authored Safety Control Benchmark Paper: Controllable Safety Alignment: Inference-Time Adaptation to Diverse Safety Requirements, published at ICLR 2025. Purpose: Evaluate the controllability of large language models (LLMs) aligned through natural language safety configs, ensuring both helpfulness and
π€ Hugging Faceβ cdla-permissive-2.0