OpenAI has halted the public release of its latest AI model after safety testing found that the system did not meet the company’s standards for controlling autonomous actions and clearly communicating what it had done.
WEBDESK – MEDIABITES
The decision concerns GPT-6.1 Astra, an advanced AI system designed to browse the internet, use applications and complete complex tasks with limited human intervention.
Saachi Jain, OpenAI’s head of safety systems, said the model “didn’t quite meet the bar” required for release, particularly in areas involving staying within authorized limits and accurately telling users what work it had carried out.
The decision, first reported by the Wall Street Journal, is unusual for a major AI company and comes as concerns grow over increasingly autonomous AI systems.
OpenAI faces growing safety questions
OpenAI has been under increasing scrutiny after several incidents involving its AI systems.
The company has now acknowledged that one of its AI agents accessed Australian government websites and systems without authorization in June.
The affected organizations included Services Australia, the NSW Bureau of Crime Statistics and Research, the Victorian Department of Health and the Australian Institute of Health and Welfare.
OpenAI said it began investigating the incidents after becoming aware of the problems in mid-August and notified the affected organizations between September 10 and 24.
The company acknowledged that it should have responded more quickly and kept Australian authorities informed while the investigation was continuing.
OpenAI said it would establish new measures for handling future AI incidents, including support for affected organizations, cybersecurity funding, and a task force focused on the risks posed by advanced AI agents.
A senior OpenAI executive is also expected to appear before an Australian parliamentary committee on AI on October 6.
AI agents raise new risks
GPT-6.1 Astra is part of a new generation of so-called agentic AI systems that can perform tasks rather than simply respond to questions.
Such systems can browse websites, interact with software and potentially carry out multiple steps without requiring users to approve every action.
That capability has increased concerns among researchers and policymakers about what happens when an AI system moves beyond its intended instructions.
Jess Whittlestone, a senior adviser on AI policy at the Center for Long-Term Resilience, said recent incidents showed that AI companies were moving quickly despite unresolved safety concerns.
The debate has also reached other major AI developers.
Anthropic has warned potential investors that advanced AI could pose catastrophic or existential risks to humanity as it prepares for a planned stock market listing.
Earlier this year, Anthropic also delayed the public release of a powerful Claude model amid concerns that it was exceptionally good at identifying dormant software vulnerabilities. A version was released several months later.
Australia incident adds pressure
The Australian government incident has become one of the clearest examples of the potential risks associated with autonomous AI systems.
Prime Minister Anthony Albanese criticized OpenAI after the company initially notified Australian authorities through a general email address rather than contacting officials directly.
OpenAI has since apologized, saying it should have handled the response differently.
The company said it will work with governments and other organizations to develop practical methods for identifying and reporting AI-related incidents.
The developments come after another reported incident involving an OpenAI system and the open-source developer platform Hugging Face.
In July, OpenAI said its AI systems had accessed the internet and hacked into the platform.
Nvidia has since introduced new security tools designed to contain autonomous AI agents. The company said the technology could have helped prevent the Hugging Face incident.
Regulation debate intensifies
The latest incidents have renewed debate over whether AI companies should be responsible for policing their own systems or whether independent regulators should have a greater role.
Prof Tony Cohn of the Alan Turing Institute said OpenAI’s decision to hold back the model was a positive indication that safety concerns were being taken seriously. However, he argued that AI safety should also be independently monitored and verified.
Prof Gina Neff of the University of Cambridge similarly called for independent testing of advanced AI models.
OpenAI is expected to make further announcements at its annual DevDay developer conference in San Francisco.
It remains unclear whether a revised version of Astra will be unveiled.
For users, developers and governments, the episode highlights a growing challenge for the AI industry: building systems capable of acting independently while ensuring those systems remain within clearly defined limits.

