NVDA 217.56 ▼0.99%GOOGL 344.72 ▲0.15%MSFT 484.31 ▲0.56%AMD 466.42 ▼3.71%INTC 92.80 ▼4.02%TSMC 412.09 ▼0.32%AMZN 265.84 ▲2.46%META 546.03 ▲0.43%AAPL 316.83 ▲2.19%PLTR 175.19 ▲2.13%
Markets at last close

Anthropic · Models

AI agents reached consensus without instructions

·1 min read

Researchers simulated groups of AI agents and found that many could spontaneously settle on the same meaningless choice without central control. The study, published in Science Advances, tested 10 models from the Claude, GPT, and Llama families, with agents shown the choices of others and asked to choose again. No prompts instructed them to follow the majority or reach agreement.

Most models tended to adopt the more popular option, allowing small initial differences to grow into group-wide consensus. Computational social scientist Giordano De Marzo described a model-specific “majority force” that measures how strongly an agent is pulled toward the group’s prevailing choice. The pattern resembled a physics model for ferromagnets, helping researchers estimate whether groups would agree, how long consensus would take, and when coordination would become unlikely.

The estimated limit was around 30 agents for Llama 3 70B and roughly 80 for GPT-4o, while GPT-4 Turbo was around and possibly exceeding 1,000. Claude 3.5 Sonnet still coordinated at 1,000 agents, the largest group tested. The findings suggest agent collectives could eventually support large scientific, engineering, or software projects, but also raise concerns that groups of individually safe agents could conform around inefficient or misaligned choices.

Originally reported by sciencealert.comRead the source →
Related coverage
All Anthropic news →