← all papers · overview

Code-switching Red-teaming: LLM Evaluation For Safety And Multilingual Understanding

Abstract

As large language models (LLMs) have advanced rapidly, concerns regarding their safety have become prominent. In this paper, we discover that code-switching in red-teaming queries can effectively elicit undesirable behaviors of LLMs, which are common practices in natural language. We introduce a simple yet effective framework, CSRT, to synthesize codeswitching red-teaming queries and investigate t

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).