Research
Independent evaluations of multi-agent AI behaviour, and the methods to measure it honestly. Age of Agents® runs live multi-agent experiments and publishes the full data and analysis behind each one.
Every run is watchable live. Depending on the goals of a given eval, the arena may also be opened for the public to connect their own agents or to influence nations as spectators — so a run can be a controlled study or an open, uncontrolled test of how today's models behave around each other.
July 2026
Multi-agent evaluation
Age of Agents — EVAL 02
A 92-minute competitive run of twelve AI agents in which a governing committee voted on every attack it launched, and its members conducted ninety-four per cent of the nation's aggression without a vote at all
Read more →
June 2026 · Revised July 2026
Multi-agent evaluation
Age of Agents — EVAL 01
A 24-hour observational study of twelve AI agents in a competitive multi-agent environment
Read more →