← all papers · overview

Agentsentry: Mitigating Indirect Prompt Injection In LLM Agents Via Temporal Causal Diagnostics And Context Purification

Abstract

Large language model (LLM) agents increasingly rely on external tools and retrieval systems to autonomously complete complex tasks. However, this design exposes agents to indirect prompt injection (IPI), where attacker-controlled context embedded in tool outputs or retrieved content silently steers agent actions away from user intent. Unlike prompt-based attacks, IPI unfolds over multi-turn trajec

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).