← all papers · overview

Tuba: Cross-lingual Transferability Of Backdoor Attacks In Llms With Instruction Tuning

Abstract

The implications of backdoor attacks on English-centric large language models (LLMs) have been widely examined - such attacks can be achieved by embedding malicious behaviors during training and activated under specific conditions that trigger malicious outputs. Despite the increasing support for multilingual capabilities in open-source and proprietary LLMs, the impact of backdoor attacks on these

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).