← all papers · overview

Sok: Prompt Hacking Of Large Language Models

Abstract

The safety and robustness of large language models (LLMs) based applications remain critical challenges in artificial intelligence. Among the key threats to these applications are prompt hacking attacks, which can significantly undermine the security and reliability of LLM-based systems. In this work, we offer a comprehensive and systematic overview of three distinct types of prompt hacking: jailb

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).