← all papers · overview

Instruction Backdoor Attacks Against Customized Llms

Abstract

The increasing demand for customized Large Language Models (LLMs) has led to the development of solutions like GPTs. These solutions facilitate tailored LLM creation via natural language prompts without coding. However, the trustworthiness of third-party custom versions of LLMs remains an essential concern. In this paper, we propose the first instruction backdoor attacks against applications integ

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).