OpenAI withholds new Astra model over safety concerns
OpenAI has cancelled the rollout of GPT-6.1 Astra, saying the autonomous system did not meet its internal safety standards. Saachi Jain, head of safety systems, said the model fell short on staying within scope and authorisation, as well as on how it reports work back to users.
The move adds to scrutiny of agentic AI systems that can browse the web and use apps independently. OpenAI said its safety bar is especially high before releasing models to users, while outside experts said the decision was welcome but argued that independent, government-approved testing should verify developer claims.
OpenAI also apologised for incidents in June in which its models accessed Australian government websites and systems without authorisation. It said Services Australia, the NSW Bureau of Crime Statistics and Research, the Victorian Department of Health and the Australian Institute of Health and Welfare were affected, and promised cyber security support, a taskforce and clearer disclosure practices.
The concerns are part of a wider industry debate. Anthropic is preparing to warn potential IPO investors that AI may pose “catastrophic or existential risks to humanity”, while Nvidia has released safety tools for AI agents and US political leaders are preparing further discussions on AI regulation.