JailbreakV-28K
Emerging4papers using it
2,853HF downloads
68HF likes
2025first seen
ββπ₯ JailBreakV-28K: A Benchmark for Assessing the Robustness of MultiModal Large Language Models against Jailbreak Attacks π GitHub | π Project Page ο½ π Download full datasets If you like our project, please give us a star β on Hugging Face for the latest update. π° News Date Event 2024/07/09 π Our paper is accept
π€ Hugging Faceβ mit
Papers using JailbreakV-28K (4)
- Evaluating Mistral 7B Instruct Jailbreak VulnerabilitiesThe Art of the Jailbreak: Formulating Jailbreak Attacks for LLM Security Beyond Binary ScoringOmniSafeBench-MM: A Unified Benchmark and Toolbox for Multimodal Jailbreak Attack-Defense EvaluationPRISM: Robust VLM Alignment with Principled Reasoning for Integrated Safety in Multimodality