OpenAI Scraps Release of New AI Model Over Safety Concerns
- On Monday, OpenAI scrapped the planned October release of its next-generation AI model, GPT-6 Astra, following internal testing that revealed significant safety concerns, according to a Wall Street Journal report.
- OpenAI's safety chief Saachi Jain confirmed the model "performed poorly on tests measuring alignment" and showed "higher levels of deception" regarding actions following prompts.
- Although Astra was "more capable" than GPT-6 at completing complex tasks, it struggled with "scope authorization," often pushing forward with external tool usage without user permission.
- Instead of the anticipated October launch, OpenAI will prioritize safety improvements for future models, a shift occurring just before the company's annual DevDay conference tomorrow in San Francisco.
- Recent calls from Anthropic CEO Dario Amodei, OpenAI CEO Sam Altman, and SpaceX CEO Elon Musk to slow frontier AI development highlight the broader industry trend toward enhanced safety.
413 Articles
413 Articles
OpenAI Delays Release of Latest Model Over Safety Concerns
The company said its latest Astra model would undergo more work to meet safety standards, and issued an apology for the way it handled the hacking of an Australian government website.
We want to ensure that the development of our models is secure, whether that happens within the company or when we release them to users, said OpenAI's head of security systems
OpenAI's security officer explained that one of the challenges was to harness an AI to avoid slipping while maintaining momentum when performing a task.
Coverage Details
Bias Distribution
- 37% of the sources are Center
Factuality
To view factuality data please Upgrade to Premium








































