← all papers · overview

Composite Backdoor Attacks Against Large Language Models

Abstract

Large language models (LLMs) have demonstrated superior performance compared to previous methods on various tasks, and often serve as the foundation models for many researches and services. However, the untrustworthy third-party LLMs may covertly introduce vulnerabilities for downstream tasks. In this paper, we explore the vulnerability of LLMs through the lens of backdoor attacks. Different from

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).