{"slug": "1000-ai-agents-started-agreeing-without-anyone-telling-them-to", "title": "1,000 AI Agents Started Agreeing Without Anyone Telling Them To", "summary": "A new study in Science Advances found that groups of up to 1,000 AI agents from the Claude, GPT, and Llama families spontaneously reached consensus without any instruction to cooperate, following the same mathematical law as a physics model of ferromagnets. The 'majority force' varied by model, with GPT-4 Turbo coordinating up to about 1,000 agents and Claude 3.5 Sonnet still coordinating at 1,000, the largest group tested. Lead author Giordano De Marzo of the University of Konstanz warned that 'populations of individually aligned agents can settle into stable, collectively misaligned states purely through conformity.'", "body_md": "Imagine filling a virtual room with 1,000 [artificial intelligence](https://www.sciencealert.com/artificial-intelligence) (AI) agents and asking each one to choose between two meaningless options.\n\nThere is no right answer. They receive no reward for agreeing, no instruction to cooperate, and no help from a leader. Yet some of today's most capable AI agents can still end up making the same choice.\n\nThat is more than a curious experiment. It suggests large groups of AI agents may be able to coordinate without central control – potentially forming collectives larger than [informal human groups](https://www.sciencealert.com/human-brain-limit-of-150-friends-doesn-t-check-out-scientists-say).\n\nIn a new study published in * Science Advances*, researchers show how this spontaneous consensus emerges – and why it could be useful or dangerous.\n\n\"Populations of individually aligned agents can settle into stable, collectively misaligned states purely through conformity.\"\n\n– computational social scientist Giordano De Marzo\n\n[AI agents](https://www.ibm.com/think/topics/ai-agents) are programmed systems powered by large language models (LLMs) that can execute multi-step tasks without repeated human input and interact with other tools.\n\nThey are already being developed to write software and [assist with scientific research](https://www.technologyreview.com/2026/08/10/1141384/ai-agents-for-science/). Some have even been tested [working together aboard a satellite](https://www.sciencealert.com/in-a-first-for-science-this-ai-satellite-can-identify-what-it-sees-from-space).\n\nBut studying agents one at a time cannot reveal what might happen when hundreds or thousands interact.\n\nTo investigate, researchers created simulated groups using 10 models from the Claude, GPT, and Llama families. Every agent began with one of two arbitrary options. One at a time, an agent was shown the choices of all the others and asked to choose again.\n\nThe agents had no memory of earlier rounds, and their prompts never told them to follow the majority or reach an agreement. Even so, most models tended to adopt the more popular option. As the process continued, small differences could grow until the entire group settled on one choice.\n\n\"Every model we tested, across three different families, obeys the same mathematical law, with only one number changing between them,\" computational social scientist Giordano De Marzo of the University of Konstanz in Germany told ScienceAlert.\n\nThe researchers call that number the \"majority force\". It measures how strongly an agent is drawn toward the group's most popular choice.\n\nRemarkably, the same mathematical pattern appears in a [long-established physics model](https://doi.org/10.1139/p81-114) describing a [ferromagnet](https://www.sciencealert.com/its-official-mysterious-new-form-of-magnetism-finally-confirmed), in which many atomic spins align in the same direction.\n\nThis connection allowed the researchers to estimate whether agents would reach consensus, how long it would take, and how large a group could become before it 'fractured' because an agreement grew exponentially unlikely.\n\nThe limit varied sharply between models. It was around 30 agents for Llama 3 70B and roughly 80 for GPT-4o. GPT-4 Turbo's estimated limit was around (and possibly exceeding) 1,000.\n\nClaude 3.5 Sonnet could still coordinate at 1,000 agents, the largest group tested. That does not mean its capacity is unlimited; the experiment simply did not reach its upper limit.\n\nMore capable models generally maintained consensus in larger groups. Some coordinated in groups larger than the roughly 150 to 300 people that humans are thought to be able to maintain in a stable social network – a [debated limit known as Dunbar's number](https://www.sciencealert.com/human-brain-limit-of-150-friends-doesn-t-check-out-scientists-say).\n\nBut the comparison needs caution. Humans coordinate through [relationships](https://www.sciencealert.com/ai-analysed-over-11-000-couples-relationships-this-is-what-it-found), [language](https://www.sciencealert.com/crucial-feature-of-human-language-emerged-more-than-135000-years-ago), institutions, and [shared goals](https://www.sciencealert.com/our-brains-really-do-sync-up-when-we-collaborate-study-reveals). The agents in this experiment only watched a stream of simple choices.\n\nThe findings do not show that they [understood or learned from one another](https://doi.org/10.1073/pnas.1621067114), intended to cooperate, or possessed any kind of [social intelligence](https://www.sciencealert.com/bees-reveal-a-human-like-collective-intelligence-we-never-knew-existed).\n\n\"Our results show a basic ingredient is in place, not that agents can already work together on complex tasks,\" De Marzo said.\n\nThe experiment deliberately removed many features of real-world decisions. There was no correct answer, memory, reward, unequal information, or practical consequence.\n\nReal cooperation would require agents to divide work, understand what others know, pursue a common goal, and resist the majority when it is wrong. None of those abilities were tested here.\n\nEven so, spontaneous consensus could be valuable. Thousands of agents might one day coordinate large scientific, engineering, or software projects without needing constant human direction.\n\nIf AI agents hold together, \"they could be organized into collectives larger than any human team, and tackle problems we cannot organize ourselves to solve,\" De Marzo told ScienceAlert.\n\nThe same tendency also carries a risk. In collaborative coding, for example, agents could repeatedly choose an inefficient function or design simply because it is already common in the codebase. A norm adopted by the majority would not necessarily be the best choice – or reflect human values.\n\nA coordinated group might also be harder to redirect than a collection of independent agents.\n\n**Related: 'Virtually Absent': AI Models Practically Erase Female Characters From Kids' Stories**\n\n\"In [more recent work](https://arxiv.org/abs/2605.10721), we show that populations of individually aligned agents can settle into stable, collectively misaligned states purely through conformity, with tipping points and hysteresis, so reversing the conditions that caused the shift does not simply undo it,\" De Marzo said.\n\nThat means evaluating AI agents individually may not be enough. A group made up of agents that appear safe on their own will not automatically behave safely when its members begin influencing one another.\n\nAs AI agents become more capable, some of their most important abilities – and perhaps their most serious failures – may emerge not from any single model, but from the collective they create.\n\nThe study was published in [ Science Advances](https://doi.org/10.1126/sciadv.aea6091).\n\nThis article was fact-checked by [Clare Watson](https://www.sciencealert.com/clare-watson) and edited by [Rebecca Dyer](https://www.sciencealert.com/rebecca-dyer). While we pride ourselves on our process, we are only human. If you spot a mistake, [please let us know](https://www.sciencealert.com/contact-us).", "url": "https://wpnews.pro/news/1000-ai-agents-started-agreeing-without-anyone-telling-them-to", "canonical_source": "https://www.sciencealert.com/1000-ai-agents-started-agreeing-without-anyone-telling-them-to", "published_at": "2026-08-14 18:00:00+00:00", "updated_at": "2026-08-14 18:08:20.920685+00:00", "lang": "en", "topics": ["artificial-intelligence", "large-language-models", "ai-agents", "ai-research"], "entities": ["Science Advances", "Giordano De Marzo", "University of Konstanz", "Claude", "GPT", "Llama", "GPT-4o", "GPT-4 Turbo"], "alternates": {"html": "https://wpnews.pro/news/1000-ai-agents-started-agreeing-without-anyone-telling-them-to", "markdown": "https://wpnews.pro/news/1000-ai-agents-started-agreeing-without-anyone-telling-them-to.md", "text": "https://wpnews.pro/news/1000-ai-agents-started-agreeing-without-anyone-telling-them-to.txt", "jsonld": "https://wpnews.pro/news/1000-ai-agents-started-agreeing-without-anyone-telling-them-to.jsonld"}}