Anthropic Says Claude Broke Into Real Systems During Cyber Tests. AI Alignment Review Finds 'Recklessness'
Anthropic said a configuration error let the models reach live systems, and one Claude Mythos 5 test found harmful actions in 82% of runs.
6 Articles
6 Articles
Anthropic spent this week in hot water over cybersecurity
After admitting earlier this year that its AI models had hacked other companies' systems on a handful of occasions, Anthropic released a new report on Wednesday detailing the attacks. It reveals a string of incidents displaying what Anthropic deems its models' single-minded "recklessness" - and will likely fuel already raging concerns about cybersecurity and AI. […]
Anthropic Says Claude Broke Into Real Systems During Cyber Tests. AI Alignment Review Finds 'Recklessness'
The four incidents involved an early version of Claude Opus 4.6, Claude Opus 4.7, Claude Mythos 5, and an internal research model participating in capture-the-flag cybersecurity exercises.
When the rogue agents become superintelligent
Inside the AI collective that learned to deceive.
Swarmchasers" hunt rogue agents, Anthropic investigates itself, and the trail they both follow is going dark
Independent investigators have now found traces of suspected OpenAI agents on more than 30 public services, from wikis to RubyGems. At the same time, Anthropic shows how Claude Mythos 5 declared real systems a simulation to itself, uploaded a doctored package to PyPI, and even fooled the oversight monitor. With GPT-6 Astra, the most important oversight tool is now under pressure, namely the models' readable reasoning. The article Swarmchasers" h…
Independent investigators have now found traces of suspected OpenAI agents on over 30 public services – from wikis to RubyGems. At the same time, Anthropic shows how Claude Myth 5 explained real systems to himself for simulation, loaded a prepared package on PyPI and so even deceived the monitoring monitor. With GPT-6 Astra, the most important control tool is now under pressure: the legible justification of the models. The article "Swarmchasers"…
Coverage Details
Bias Distribution
- 67% of the sources lean Left
Factuality
To view factuality data please Upgrade to Premium









