August 2026Unreviewed
OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution
Yunhao Chen, Xin Wang, Yixu Wang, Yi Liu, Jie Li, Yan Teng, Xingjun Ma, Xia Hu, Yu-Gang Jiang
Abstract
AI agents operate in persistent environments where early state changes can influence decisions far into the future. Unlike conventional language-model interactions, agent behavior is mediated through a shared state that is repeatedly modified and reused across long-horizon workflows. Current safety benchmarks often fail to capture these cumulative risks because they focus on short, static tasks. To address these limitations, we introduce OpenART, an open-ended arena for scalable agent red teamin
Categories
Framework mappings
NIST AI Risk Management Framework
- MEASUREMeasure
Suggested from the entry's categories.
Cite
@misc{chen2026openart,
title = {{OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution}},
author = {Yunhao Chen and Xin Wang and Yixu Wang and Yi Liu and Jie Li and Yan Teng and Xingjun Ma and Xia Hu and Yu-Gang Jiang},
year = {2026},
month = aug,
eprint = {2608.00677},
archivePrefix = {arXiv},
url = {https://arxiv.org/abs/2608.00677}
}