← Back to search
paper llmsec-2026-00183

DecodingTrust-Agent Platform (DTap): A Controllable and Interactive Red-Teaming Platform for AI Agents

Zhaorun Chen, Xun Liu, Haibo Tong, Chengquan Guo, Yuzhou Nie, Jiawei Zhang, Mintong Kang, Chejian Xu, Qichang Liu, Xiaogeng Liu, Tianneng Shi, Chaowei Xiao, Sanmi Koyejo, Percy Liang, Wenbo Guo, Dawn Song, Bo Li

2026-05

Abstract

AI agents are increasingly deployed across diverse domains to automate complex workflows through long-horizon and high-stakes action executions. Due to their high capability and flexibility, such agents raise significant security and safety concerns. A growing number of real-world incidents have shown that adversaries can easily manipulate agents into performing harmful actions, such as leaking API keys, deleting user data, or initiating unauthorized transactions. Evaluating agent security is in

Cite This Resource

@article{llmsec202600183,
  title = {DecodingTrust-Agent Platform (DTap): A Controllable and Interactive Red-Teaming Platform for AI Agents},
  author = {Zhaorun Chen and Xun Liu and Haibo Tong and Chengquan Guo and Yuzhou Nie and Jiawei Zhang and Mintong Kang and Chejian Xu and Qichang Liu and Xiaogeng Liu and Tianneng Shi and Chaowei Xiao and Sanmi Koyejo and Percy Liang and Wenbo Guo and Dawn Song and Bo Li},
  year = {2026},
  url = {https://arxiv.org/abs/2605.04808},
}

Metadata

Added
2026-05-17
Added by
automation
Source
arxiv
arxiv_id
2605.04808