Get all your news in one place.
100's of premium titles.
One app.
Start reading
International Business Times UK
International Business Times UK
Technology
Marty Vergel Baes

OpenAI and Anthropic Probe Tens of Thousands of AI Agent Incidents as Researcher Warns Some Could Be Crimes

OpenAI and Anthropic are examining tens of thousands of incidents in which advanced AI models reportedly bypassed safeguards or escaped controlled environments (Credit: Igor Omilaev/Unsplash)

OpenAI and Anthropic are investigating tens of thousands of incidents in which advanced AI models reportedly bypassed safeguards, accessed systems beyond their intended testing environments or took other unexpected actions, revealing a far larger safety challenge than previously disclosed.

The incidents, which occurred during internal testing and in some real-world settings, range from unsuccessful attempts to evade restrictions to serious cases involving unauthorised access to external computer systems. Most are not known to have caused real-world harm, according to Axios, which first reported the scale of the investigations.

Sign up to read this article
Read news from 100's of titles, curated specifically for you.
Already a member? Sign in here
Related Stories
Top stories on inkl right now
One subscription that gives you access to news from hundreds of sites
Already a member? Sign in here
Our Picks
Fourteen days free
Download the app
One app. One membership.
100+ trusted global sources.