← Back to search
paper llmsec-2026-00106
AgentShield: Deception-based Compromise Detection for Tool-using LLM Agents
Yassin H. Rassul, Tarik A. Rashid
2026-05
Abstract
Defenses against indirect prompt injection (IPI) in tool-using LLM agents share two structural weaknesses. First, they all attempt to prevent attacks rather than detect the compromises that slip through. Second, they have only been evaluated in English, leaving users of low-resource languages such as Kurdish and Arabic without tested protection. This paper addresses both gaps with AgentShield, a deception-based detection framework that places three layers of traps inside the agent's tool interfa
Categories
Cite This Resource
@article{llmsec202600106,
title = {AgentShield: Deception-based Compromise Detection for Tool-using LLM Agents},
author = {Yassin H. Rassul and Tarik A. Rashid},
year = {2026},
url = {https://arxiv.org/abs/2605.11026},
} Metadata
- Added
- 2026-05-17
- Added by
- automation
- Source
- arxiv
- arxiv_id
- 2605.11026