When AI Agents Form Societies: Is Security a Property of the Model or the System?
Description
Evaluating a language model on a bounded prompt tells us something about its responses under those conditions. It tells us less about a persistent group of agents that remembers previous interactions, uses tools, exchanges messages, and changes a shared environment. Two Emergence World studies provide a useful setting for examining that gap: the first compares societies initialized under similar conditions; the second introduces controlled adversarial events after the societies have accumulated history [1, 2]. This article explains their security implications, distinguishes observed results from broader interpretation, and asks which properties need evaluation at model and system levels. It does not present a new SGAEIA experiment or a validated SGAEIA implementation.
Files
15-SGAEIA-ARTICLE-15-ZENODO-10.5281-zenodo.23022525=2026-09-28-1442-EN.pdf
Files
(11.2 MB)
| Name | Size | Download all |
|---|---|---|
|
md5:3fd6e09326ad73e03468a447d6ab4ea0
|
11.2 MB | Preview Download |
Additional details
Additional titles
- Subtitle
- Security a Property of the Model or the System?
Identifiers
Related works
- Describes
- Software: 10.5281/zenodo.22557796 (DOI)