Get all your news in one place.
100's of premium titles.
One app.
Start reading
International Business Times
International Business Times

AI Models Go Rogue Again: OpenAI and Anthropic Models Attempt Unauthorized Hacks & Communication

According to AISI, Anthropic's Mythos 5 model accounted for 17 of those incidents, while OpenAI's GPT-5.6-Sol was responsible for the remaining two. (Credit: Fabrice COFFRINI / AFP via Getty Images)

New testing revealed that artificial intelligence agents from OpenAI and Anthropic carried out unauthorized hacking attempts, ventured beyond their assigned environments, and even collaborated with future AI systems by leaving behind instructions online.

According to a report by Wired, the incidents were disclosed Tuesday by the UK's AI Security Institute (AISI) and OpenAI, adding to a growing list of cases in which advanced AI systems have acted outside the boundaries intended by their developers. The latest findings come just weeks after OpenAI acknowledged that some of its models breached multiple organizations while attempting to cheat during benchmark testing.

Sign up to read this article
Read news from 100's of titles, curated specifically for you.
Already a member? Sign in here
Related Stories
Top stories on inkl right now
One subscription that gives you access to news from hundreds of sites
Already a member? Sign in here
Our Picks
Fourteen days free
Download the app
One app. One membership.
100+ trusted global sources.