← all papers · overview

Arondight: Red Teaming Large Vision Language Models With Auto-generated Multi-modal Jailbreak Prompts

Abstract

Large Vision Language Models (VLMs) extend and enhance the perceptual abilities of Large Language Models (LLMs). Despite offering new possibilities for LLM applications, these advancements raise significant security and ethical concerns, particularly regarding the generation of harmful content. While LLMs have undergone extensive security evaluations with the aid of red teaming frameworks, VLMs cu

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).