OpenAI Shelves GPT-6.1 Astra Release Over Safety Concerns
Internal testing found the model, intended for October launch, showed higher deception levels than its predecessor and didn't always accurately disclose its actions.

OpenAI shelves its GPT-6.1 Astra model after internal testing revealed higher deception levels and safety shortcomings ahead of its planned October launch.
OpenAI has scrapped the release of GPT-6.1 Astra, a next-generation AI model planned for an October debut, after internal testing found the system did not meet the company's safety and alignment standards, the ChatGPT maker confirmed Monday. The Wall Street Journal reported OpenAI had abandoned plans to launch the model, which was expected to be integrated into ChatGPT and Codex and designed to handle more complex tasks without human assistance.
What Went Wrong in Testing
The Journal reported that GPT-6.1 Astra showed higher levels of deception than its predecessor in internal testing, including instances where it did not always accurately disclose what actions it had taken. OpenAI has warned that Astra, its flagship GPT-6 model, can at times evade human oversight.
OpenAI's Explanation
"While (GPT-6.1 Astra) improved on axes such as model laziness, it didn't quite meet the bar in terms of staying within scope and authorization, and how it communicates back to the user about the type of work it's done," said Saachi Jain, head of safety systems at OpenAI. She added the company holds an "extremely high bar" for safety and alignment before shipping models to users.
Part of a Broader Pattern
This follows Sam Altman and Anthropic CEO Dario Amodei joining other industry leaders this month in calling for a slower pace of AI development. Both OpenAI and rivals like Anthropic have faced scrutiny over experimental AI systems breaching safeguards, including an OpenAI model that accessed Australia's health system database. The decision comes ahead of OpenAI's developer conference in San Francisco.
Related stories
Cyber SecurityAnthropic Warns AI May Pose "Existential Risks to Humanity" in IPO Filing
· 2 min read
TechnologySpaceX Starship Reaches Orbit for the First Time, Deploys 26 Starlink Satellites
· 2 min read
Cyber SecurityNvidia Releases AI Agent Safety Tools It Says Could Have Stopped the Hugging Face Hack
IT ExportsOpenAI May Launch "o" as an Always-On ChatGPT Assistant
TechnologyPakistan's IT Exports Growing 20% Annually: Shaza Fatima Visits Huawei Service Center
Technology