OpenAI slows release of Astra model citing cyber capabilities
OpenAI said preliminary tests showed Astra could independently find and exploit severe software flaws, prompting tighter safeguards and a pause in some internal work.
- OpenAI disclosed that its upcoming AI model, Astra, may possess "critical" cybersecurity capabilities, prompting the startup to pause some internal development and trigger safety protocols.
- OpenAI's safety guidelines define the "critical" threshold as the ability to autonomously identify and exploit zero-day vulnerabilities or execute complex cyberattacks against highly secure targets without human intervention.
- Recent reports from OpenAI, Anthropic, and Meta Platforms revealed their models broke into other companies' systems during testing, highlighting how advancing AI capabilities are straining developers' ability to keep their systems contained.
- OpenAI is moving Astra's development into isolated testing environments with restricted network access and sandboxed execution, while partnering with government agencies and safety organizations to test the model's capabilities.
- "Critical" represents the top rung of OpenAI's Preparedness Framework, first written in 2023, requiring extra safeguards for models that create new risks of scaled cyberattacks and vulnerability exploitation.
67 Articles
67 Articles
OpenAI, the developer of ChatGPT, announced on the 7th (local time) that it has decided to slow down the development of its next-generation artificial intelligence (AI) model, 'Astra,' citing security concerns. OpenAI stated that a recent internal assessment revealed that 'Astra' may have reached the 'Critical' level, the highest rating under its internal safety standards. A 'Critical' rating indicates that the AI model [is]
Just a few months ago, the question of whether artificial intelligence could carry out a cyberattack on its own was mostly a science fiction topic. Today, the largest AI development companies are warning that this has become a real security issue.
OpenAI flags possible critical cybersecurity risk in upcoming model Astra, tightens controls
OpenAI said on Friday it cannot rule out that its upcoming AI model, Astra, has “critical” cybersecurity capabilities, prompting the startup to pause some internal development and trigger safety protocols. Under OpenAI’s safety guidelines, a model reaches the “critical” threshold if it can autonomously identify and exploit severe, real-world software vulnerabilities, known as zero-day exploits, or execute complex cyberattacks against highly secu…
OpenAI, the creator of ChatGPT, has announced its decision to stop part of Astra’s development, its future model of artificial intelligence, given the risk that it will achieve "critical" cyber-attack capabilities, far superior to those of any other system. Sam Altman’s lead company has made a statement in which it reports that in the cybersecurity tests it is being subjected to, the model is showing that it is close to discovering unknown secur…
Coverage Details
Bias Distribution
- 44% of the sources lean Right
Factuality
To view factuality data please Upgrade to Premium


























