← all papers · overview

The Illusion Of Role Separation: Hidden Shortcuts In LLM Role Learning (and How To Fix Them)

Abstract

Large language models (LLMs) that integrate multiple input roles (e.g., system instructions, user queries, external tool outputs) are increasingly prevalent in practice. Ensuring that the model accurately distinguishes messages from each role -- a concept we call *role separation* -- is crucial for consistent multi-role behavior. Although recent work often targets state-of-the-art prompt injection

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).