← all papers · overview

Cleangen: Mitigating Backdoor Attacks For Generation Tasks In Large Language Models

Abstract

The remarkable performance of large language models (LLMs) in generation tasks has enabled practitioners to leverage publicly available models to power custom applications, such as chatbots and virtual assistants. However, the data used to train or fine-tune these LLMs is often undisclosed, allowing an attacker to compromise the data and inject backdoors into the models. In this paper, we develop

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).