OpenAI will not release its next‑generation model - GPT‑6.1 Astra - due to safety concerns, the ChatGPT‑maker confirmed on Tuesday.
The AI system — which performs tasks like browsing the web and using apps by itself — did not meet the company’s internal safety bar, Saachi Jain, head of safety systems at OpenAI, said.
In recent weeks, top AI leaders, including OpenAI’s Sam Altman and Anthropic boss Dario Amodei, have urged the industry to slow its pace of development because of growing risks.
OpenAI’s decision to pull the model is a rare instance of a major AI developer retracting a rollout over safety concerns, first reported by the Wall Street Journal.
Critics say the model fell short in staying within scope and authorization, and in how it communicates back to the user about the work it has completed.
Jain added, “We want to make sure our model development is safe no matter whether that’s in the company, or when we ship it to users. But when we ship it to users, we have an extremely high bar in terms of safety and alignment.”
The flagship GPT‑6 Astra agentic model, released in September, specializes in complex reasoning and autonomous task execution and was touted as the result of “years of research and big bets.”
Security controls have come under intense scrutiny after several high‑profile incidents involving OpenAI’s technology.
Last week, Australian Prime Minister Anthony Albanese announced that a rogue OpenAI agent had hacked into a government website in June and accessed private data in what experts said was the first known case of its kind in the world.
In July, OpenAI said its AI systems had accessed the internet and hacked into the open‑source developer hub Hugging Face, prompting calls for tighter controls over the technology.
On Monday, AI chip giant Nvidia released a set of software safety tools for autonomous AI platforms — called agents — that it said could have prevented the Hugging Face hack.
One of the new tools uses hardware features in Nvidia’s chips to contain agents. Nvidia boss Jensen Huang has largely dismissed calls for tighter AI regulations, arguing that rogue agents are an engineering problem that can be solved.
Nvidia also agreed to buy Hugging Face for $12.9 bn (£9.74 bn) earlier this month.















