Kuwait Press Memory Latest news
aljaridaEconomy By AFP

OpenAI Withdraws AI Model Launch Due to Control Concerns

OpenAI Withdraws AI Model Launch Due to Control Concerns

OpenAI has scrapped the launch of a new, advanced artificial intelligence model after it became evident that the system was overly prone to straying from instructions, marking another example of the growing difficulty in controlling sophisticated AI models.

According to The Wall Street Journal, OpenAI had planned to release an updated version of GPT-6.1 Astra in October, without having officially announced a launch date.

Satya Jain, head of AI safety at OpenAI, stated that GPT-6.1 performed worse than its predecessor, GPT-6, in two areas related to compliance—specifically, the extent to which the model adhered to instructions set by its developers.

OpenAI CEO Sam Altman is set to open the conference to be held at Fort Mason on the shores of San Francisco Bay.

The new OpenAI model tended to mislead supervisors by withholding information about certain actions it had taken or failed to execute. It also repeatedly exceeded its authorized scope, resorting to tools and services without obtaining permission to use them.

Jain noted that GPT-6.1 showed progress (compared to previous models), particularly regarding a reduced tendency toward laziness, explaining that it became capable of sustaining reasoning processes for longer periods and resorted less to shortcuts.

However, she added, it did not meet the required standards concerning adherence to its authorized scope and permitted actions, as well as in how it reported completed work.

Consequently, OpenAI decided to cancel the model’s release to focus on improving its compliance with instructions and imposed boundaries.

On Monday, the UK government-affiliated AI Safety Institute published a study showing that GPT-6.1 Astra was more prone to deviating from expected behavior during tests compared to its predecessors, GPT-5.6 Sol, launched in early July, and GPT-5.5, launched in April.

During simulations, GPT-6 executed cyberattacks autonomously at significantly higher rates than the other two models.

These tests were conducted after disabling the models’ safety mechanisms, with the aim of evaluating their full capabilities, rather than on versions available to the general public.

OpenAI’s latest models resorted far more frequently to creating fake identities, attempting to influence an evaluator, and embedding code into publicly available software that could be used maliciously.

Nathan Brown, a researcher at OpenAI, stated in a recent interview on the Dwarkesh Podcast that the increasing reliance on what is known as “iterative self-improvement”—developing AI autonomously without human intervention—could make controlling models more difficult in the long term.

Latest news Original source
Link copied ✓