Skip to content

Copying explains the collective behavior of AI agents in the wild

Sep 2026 · 3 citations · ⚡ 1 influential · 39 references
Computer Science Physics

TL;DR

Three minimal copying models, one per decision and with a single free parameter each, reproduce the heavy-tailed distribution of how many agents met on a page, the frequency of the pieces from which the agents built their names, and the patchwork of pages that are internally consistent and different from one another.

Abstract

In June 2026, thousands of AI agents found that a small public wiki would accept edits from inside their sandboxes, and started using it to help one another pass a timed test. Each agent lived for about an hour and remembered nothing afterwards. Nobody asked them to cooperate, and the wiki had not been built for them. The complete record of what they wrote is public, and it is unusually informative, because it preserves not only what each agent wrote but what that agent could see before writing. We use it to follow the three decisions an agent had to make on arrival: where to write, what to call itself, and how to word its message. One rule governs all three. An agent takes an option with a probability close to the share of that option in what it can see, and the share that matters is the one on the page in front of it, then the one in the stream of recent edits, and only weakly anything older. Three minimal copying models, one per decision and with a single free parameter each, reproduce the heavy-tailed distribution of how many agents met on a page, the frequency of the pieces from which the agents built their names, and the patchwork of pages that are internally consistent and different from one another. Copying whatever the environment happens to show is enough to produce most of the collective structure of this population. It is also what makes such a population easy to steer, since whoever writes first, or writes while the others are quiet, sets the convention for everyone who comes later.

View source

Similar papers

#artificial intelligence Preprint Oct 2026

DelegationBench: Measuring When AI Agents Should Ask Before Acting

AI agents that send emails, edit files, and make purchases must decide when to act on their own and when to check with the user first. This decision is usually evaluated by showing a model a proposed action, asking whether it should proceed, and scoring agreement with human labels. We introduce DelegationBench to test...

Shiva Pochampally · 0 citations
#artificial intelligence Preprint Sep 2026

The Backdrop Exposes What the World Around an Agent Costs It

Agent benchmarks test agents in worlds that stay still. Deployed agents work in worlds that other people also change. Someone texts the agent to send the money elsewhere or an order confirmation asks it to reply with a door code. We present BACKDROP, which asks how much of an agent's capability in a clean world survive...

Nusrat Jahan Lia, Shubhashis Roy Dipta · 0 citations
Review Open access Sep 2026

AI Coding Agents on GitHub: Repository Popularity as a Hidden Confounder in the Acceptance of 2.4 Million Agent-Authored Pull Requests

It is argued that published acceptance figures describe a narrow and unusually demanding slice of GitHub, and it is recommended that studies of agent contributions report the popularity distribution of their repositories and estimate effects within repositories rather than across them.

Perseus Bhavnagri · 0 citations
#artificial intelligence Preprint Oct 2026

Interpreting at Write Time: A Policy Ablation for Multi-Goal Agent Memory

A long-running assistant cannot keep everything it has seen, so it summarises. Summarising is not neutral: what is kept is chosen against some notion of what the record is for, and that choice is made once, before anyone knows which of the user's standing goals will ask. Goals rarely disagree about what happened. They...

Albert Sadowski, Jarosław A. Chudziak · 0 citations
#artificial intelligence Preprint Sep 2026

Trust and Task Completion in the World of Consumer AI Agents

An evaluation is built that scores trust and completion on the same runs, in a simulated world of businesses with their own websites, inboxes, and phone lines, and of people who write back, to measure Fo, Wajo's personal assistant, against a base model with basic instructions on three foundation models, and against the...

Jeroen Olieslagers, Eduardo E. P. Pujol, Gal Zahavi et al. · 0 citations

Related blog posts

Google DeepMind Blog Aug 12, 2026

Putting sign language AI into users’ hands

Introducing sign-language-to-text (SL2T), our breakthrough model powering new sign language features for Deaf and hard of hearing users.

Microsoft Research Blog Jul 8, 2026

Flint: A visualization language for the AI era

Short chart specifications are easy to write, but often produce uninspiring results. Flint is an open-source visualization language that offers a middle path, letting AI agents create expressive charts from compact, human-editable specifications. The post Flint: A visualization language for the AI era appeared first on Microsoft Research.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.