← all papers · overview

Target-driven Attack For Large Language Models

Abstract

Current large language models (LLM) provide a strong foundation for large-scale user-oriented natural language tasks. Many users can easily inject adversarial text or instructions through the user interface, thus causing LLM model security challenges like the language model not giving the correct answer. Although there is currently a large amount of research on black-box attacks, most of these bla

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).