← Back to search
paper llmsec-2026-00107
Position: AI Security Policy Should Target Systems, Not Models
Michael A. Riegler, Inga Strümke
2026-05
Abstract
We present swarm-attack, an open-source adversarial testing framework in which multiple lightweight LLM agents coordinate through shared memory, parallel exploration, and evolutionary optimization. Together, our results demonstrate that both safety bypass of frontier models and software vulnerability discovery, i.e., the capability class that motivated restricted release of Anthropic's Mythos Preview, are achievable at effectively zero cost using commodity hardware and openly available models. W
Categories
Cite This Resource
@article{llmsec202600107,
title = {Position: AI Security Policy Should Target Systems, Not Models},
author = {Michael A. Riegler and Inga Strümke},
year = {2026},
url = {https://arxiv.org/abs/2605.09504},
} Metadata
- Added
- 2026-05-17
- Added by
- automation
- Source
- arxiv
- arxiv_id
- 2605.09504