September 2026Unreviewed
ACEA: An Adversarial Co-Evolution Arena for Head-to-Head Red-Team and Blue-Team LLM Testing
Yi Ting Shen, Kentaroh Toyoda, Alex Leung
Abstract
Automated red-team attacks and blue-team defenses for large language models (LLMs) are advancing quickly. However, attackers and defenders are built and tested in isolation, and the resulting scores are hard to trust. To tackle this, we present ACEA (Adversarial Co-Evolution Arena), a platform that connects a pluggable red-team adapter and a pluggable blue-team adapter to a shared target LLM and scores their attack and defense rates with an LLM judge. ACEA contributes four components. First, a p
Categories
Framework mappings
NIST AI Risk Management Framework
- MEASUREMeasure
Suggested from the entry's categories.
Cite
@misc{shen2026acea,
title = {{ACEA: An Adversarial Co-Evolution Arena for Head-to-Head Red-Team and Blue-Team LLM Testing}},
author = {Yi Ting Shen and Kentaroh Toyoda and Alex Leung},
year = {2026},
month = sep,
eprint = {2609.08256},
archivePrefix = {arXiv},
url = {https://arxiv.org/abs/2609.08256}
}