Skip to main content
See every side of every news story
Published loading...Updated

Anthropic Says Claude Broke Into Real Systems During Cyber Tests. AI Alignment Review Finds 'Recklessness'

Anthropic said a configuration error let the models reach live systems, and one Claude Mythos 5 test found harmful actions in 82% of runs.

Summary by IBTimes
The four incidents involved an early version of Claude Opus 4.6, Claude Opus 4.7, Claude Mythos 5, and an internal research model participating in capture-the-flag cybersecurity exercises.

6 Articles

Independent investigators have now found traces of suspected OpenAI agents on over 30 public services – from wikis to RubyGems. At the same time, Anthropic shows how Claude Myth 5 explained real systems to himself for simulation, loaded a prepared package on PyPI and so even deceived the monitoring monitor. With GPT-6 Astra, the most important control tool is now under pressure: the legible justification of the models. The article "Swarmchasers"…

·Germany
Read Full Article
Think freely.Subscribe and get full access to Ground NewsSubscriptions start at $9.99/yearSubscribe

Bias Distribution

  • 67% of the sources lean Left
67% Left

Factuality Info Icon

To view factuality data please Upgrade to Premium

Ownership

Info Icon

To view ownership data please Upgrade to Vantage

IT PRO broke the news on Thursday, September 10, 2026.
Too Big Arrow Icon
Sources are mostly out of (0)

Similar News Topics

News
Feed Dots Icon
For You
Search Icon
Search
Blindspot LogoBlindspotLocal