← Back to search
paper llmsec-2026-00106

AgentShield: Deception-based Compromise Detection for Tool-using LLM Agents

Yassin H. Rassul, Tarik A. Rashid

2026-05

Abstract

Defenses against indirect prompt injection (IPI) in tool-using LLM agents share two structural weaknesses. First, they all attempt to prevent attacks rather than detect the compromises that slip through. Second, they have only been evaluated in English, leaving users of low-resource languages such as Kurdish and Arabic without tested protection. This paper addresses both gaps with AgentShield, a deception-based detection framework that places three layers of traps inside the agent's tool interfa

Cite This Resource

@article{llmsec202600106,
  title = {AgentShield: Deception-based Compromise Detection for Tool-using LLM Agents},
  author = {Yassin H. Rassul and Tarik A. Rashid},
  year = {2026},
  url = {https://arxiv.org/abs/2605.11026},
}

Metadata

Added
2026-05-17
Added by
automation
Source
arxiv
arxiv_id
2605.11026