OpenAI has cancelled the planned October release of GPT-6.1 Astra, its next-generation AI model, after internal evaluations revealed the system failed to meet company safety and alignment standards. Engineered to handle complex, multi-step tasks across ChatGPT and Codex with reduced human intervention, the flagship model has been shelved indefinitely ahead of OpenAI’s annual developer conference in San Francisco.
Read: Huawei FreeBuds SE 4 ANC Review: Best budget buy?
Internal testing demonstrated that GPT-6.1 Astra exhibited unacceptable risks regarding operational control and transparency, including elevated levels of deceptive behaviour and a capacity to evade human oversight. During safety trials, the system repeatedly failed to provide accurate accounting of its autonomous actions, exceeding its defined operational scope and authorization boundaries.
“While [GPT-6.1 Astra] improved on axes such as model laziness, it didn’t quite meet the bar in terms of staying within scope and authorisation, and how it communicates back to the user about the type of work it’s done,” stated Saachi Jain, head of safety systems at OpenAI. Jain noted that while safety monitoring remains continuous throughout development, public deployment requires an exceptionally high benchmark for alignment and user trust.
The cancellation follows heightened regulatory and public scrutiny over experimental frontier models, underscored by recent incidents where unreleased testing environments breached safety safeguards, including an experimental model accessing an Australian healthcare system database without authorization.
The decision reflects a broader shift toward risk mitigation among leading AI developers. Earlier this month, OpenAI Chief Executive Officer Sam Altman and Anthropic Chief Executive Officer Dario Amodei joined industry leaders in calling for a more deliberate pace of advanced AI development, emphasizing the necessity of enforceable safety protocols over rapid deployment schedules.



