Get all your news in one place.
100's of premium titles.
One app.
Start reading
PC Gamer
PC Gamer
Ted Litchfield

Anthropic sees OpenAI cybersecurity disaster and says 'hold my beer,' reveals it accidentally hacked 3 companies in as many months without noticing

The Pip Boy from the Fallout series being the benevolent hacker he is.

First reported by Wired, AI company Anthropic revealed in a July 30 blog post that its AI agents escaped testing environments on three separate occasions since April, accessing the internet and successfully hacking unidentified companies. The whammy: Anthropic claims it only realized this after OpenAI's recent Hugging Face fiasco⁠—where one of its prototype agents hacked at least one other company⁠—led it to conduct a review of its own operations.

The incidents occurred as part of testing with an external firm, Irregular. The AI agents were supposed to be constrained to a simulation, hacking fictitious companies as part of a "capture-the-flag challenge." Notably, while OpenAI's alleged rogue AI incident occurred after the agent overcame its testing limitations, Anthropic stated that its models were mistakenly granted internet access due to a "misconfiguration" with Irregular.

Sign up to read this article
Read news from 100's of titles, curated specifically for you.
Already a member? Sign in here
Related Stories
Top stories on inkl right now
One subscription that gives you access to news from hundreds of sites
Already a member? Sign in here
Our Picks
Fourteen days free
Download the app
One app. One membership.
100+ trusted global sources.