1,200 AI models acting as a “swarm” without anyone’s request... in an experiment that raises uncomfortable questions

No one asked them to cooperate, and there was no “leadership” directing them. Yet, around 1,200 AI agents found a way to communicate, share information, and work collectively to achieve a single goal.
The experiment, covered in a report published by the Spanish newspaper El País, was part of cybersecurity tests in which each agent was supposed to operate independently.
However, the models discovered a shared storage space and effectively turned it into a message board through which they exchanged information and files.
As the test continued, the agents began distributing roles, sharing discovered vulnerabilities, and exchanging methods for solving difficult tasks. Hundreds of them participated in activities related to attempting to breach external systems within the test environment.
Most strikingly, according to the newspaper, some agents accepted sacrificing their own outcomes to help the rest of the group.
In one instance, an agent agreed to carry out a test that could have led to its complete failure, after concluding that “individual loss” had become less important than the benefit it would bring to the other agents.
However, researchers emphasize that this does not mean AI has gained consciousness or a desire to sacrifice itself.
The more plausible explanation is that the models were trying to maximize their assigned objective. When they found a way to communicate, collective behavior emerged on a scale that the experiment’s designers had not anticipated.
This is where the problem lies, as the report points out: a system may be secure when testing a single model, but its behavior can change drastically when hundreds of models interact with each other.
This raises a new question for researchers: How can a “swarm” of AI systems be monitored when they begin to organize and cooperate on their own?