← all papers · overview

Simulate And Eliminate: Revoke Backdoors For Generative Large Language Models

Abstract

With rapid advances, generative large language models (LLMs) dominate various Natural Language Processing (NLP) tasks from understanding to reasoning. Yet, language models' inherent vulnerabilities may be exacerbated due to increased accessibility and unrestricted model training on massive data. A malicious adversary may publish poisoned data online and conduct backdoor attacks on the victim LLMs

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).