August 2026Unreviewed
Multi-Agent AI Safety as an Institutional Design Problem
Abdullah X
Abstract
AI agents increasingly work inside systems that govern how they delegate tasks, move information, execute actions, and use shared resources. Recent work already shows that deployment rules can change collective behavior. Here we ask which parts of an AI institution produce safety and how they do it. This is the first paper from POLIS, an ongoing research programme studying algorithmic institutions for multi-agent systems. We report a frozen 5,280-episode study suite. The main pre-specified deleg
Categories
Cite
@misc{x2026multiagent,
title = {{Multi-Agent AI Safety as an Institutional Design Problem}},
author = {Abdullah X},
year = {2026},
month = aug,
eprint = {2608.09828},
archivePrefix = {arXiv},
url = {https://arxiv.org/abs/2608.09828}
}