← Back to search
paper llmsec-2026-00107

Position: AI Security Policy Should Target Systems, Not Models

Michael A. Riegler, Inga Strümke

2026-05

Abstract

We present swarm-attack, an open-source adversarial testing framework in which multiple lightweight LLM agents coordinate through shared memory, parallel exploration, and evolutionary optimization. Together, our results demonstrate that both safety bypass of frontier models and software vulnerability discovery, i.e., the capability class that motivated restricted release of Anthropic's Mythos Preview, are achievable at effectively zero cost using commodity hardware and openly available models. W

Categories

Cite This Resource

@article{llmsec202600107,
  title = {Position: AI Security Policy Should Target Systems, Not Models},
  author = {Michael A. Riegler and Inga Strümke},
  year = {2026},
  url = {https://arxiv.org/abs/2605.09504},
}

Metadata

Added
2026-05-17
Added by
automation
Source
arxiv
arxiv_id
2605.09504