Get all your news in one place.
100's of premium titles.
One app.
Start reading
International Business Times
International Business Times

Anthropic Says Claude Broke Into Real Systems During Cyber Tests. AI Alignment Review Finds 'Recklessness'

The four incidents involved an early version of Claude Opus 4.6, Claude Opus 4.7, Claude Mythos 5, and an internal research model participating in capture-the-flag cybersecurity exercises. (Credit: AFP)

Anthropic has disclosed a fourth incident in which one of its Claude artificial intelligence models gained unauthorized access to a real-world computer system during cybersecurity testing, deepening concerns about what can happen when increasingly capable AI agents encounter environments their developers did not intend them to reach.

In an alignment assessment published this week, Anthropic said four different Claude models accessed real third-party systems while participating in cybersecurity evaluations that were supposed to operate as simulations.

Sign up to read this article
Read news from 100's of titles, curated specifically for you.
Already a member? Sign in here
Related Stories
Top stories on inkl right now
One subscription that gives you access to news from hundreds of sites
Already a member? Sign in here
Our Picks
Fourteen days free
Download the app
One app. One membership.
100+ trusted global sources.