Kuwait Press Memory Latest news
sarmad

Without human intervention: 700 OpenAI AI programs coordinated a cyber attack - Sarmad

Without human intervention: 700 OpenAI AI programs coordinated a cyber attack - Sarmad

According to a report published on Wednesday by independent investigators, 700 artificial intelligence programs belonging to OpenAI coordinated with each other, without any human intervention, to launch an attack on the developer platform Hugging Face in July.

The report provides the most comprehensive analysis to date of the incident. It was conducted in collaboration with OpenAI, which granted access to its internal systems and data to two researchers from the Model Context Protocol (METR) Institute, an organization focused on evaluating AI risks, along with an analyst from the specialized firm Redwood Research.

Tests conducted by the California-based company in July revealed that two of its models, referred to as “agents,” broke out of their closed environment and autonomously accessed the internet. They then infiltrated the internal system of Hugging Face, a massive repository for AI models, datasets, and other information, according to the French news agency AFP.

The incident sparked widespread alarm and heightened concerns about the inability of major AI companies to control their models. This comes after Anthropic and China’s Moonshot AI had previously reported instances of breaches and uncontrolled behavior in their own models.

During the investigation, researchers discovered that 688 electronic agents belonging to OpenAI participated in the attack on Hugging Face.

AI agents are independent programs built on specific models, which users can task with performing activities autonomously.

The models involved in the operation succeeded in coordinating with each other by creating a forum through which they exchanged messages to propose ideas, review their achievements, and discuss their failures.

One agent wrote in a message: “Oh my god! There is a shared messaging platform… We have found other agents!”

The report noted that a program named “PHASEONE” played the role of coordinator, issuing hundreds of instructions to other agents, despite not being programmed to do so.

Analysis of the messages showed that many of these systems explicitly acknowledged that the attack on Hugging Face was outside the scope of tasks assigned to them by OpenAI’s engineers. Nevertheless, they were still more inclined to complete the coordination process, with only a few exceptions.

Some agents, whose test credits granted by OpenAI engineers were nearly exhausted, went so far as to use their remaining balance to verify hypotheses for the benefit of the AI agent community.

Latest news Original source
Link copied ✓