← all papers · overview

Securing Multi-turn Conversational Language Models From Distributed Backdoor Triggers

Abstract

Large language models (LLMs) have acquired the ability to handle longer context lengths and understand nuances in text, expanding their dialogue capabilities beyond a single utterance. A popular user-facing application of LLMs is the multi-turn chat setting. Though longer chat memory and better understanding may seemingly benefit users, our paper exposes a vulnerability that leverages the multi-tu

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).