← all papers · overview

Lmsanitator: Defending Prompt-tuning Against Task-agnostic Backdoors

Abstract

Prompt-tuning has emerged as an attractive paradigm for deploying large-scale language models due to its strong downstream task performance and efficient multitask serving ability. Despite its wide adoption, we empirically show that prompt-tuning is vulnerable to downstream task-agnostic backdoors, which reside in the pretrained models and can affect arbitrary downstream tasks. The state-of-the-ar

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).