Get all your news in one place.
100's of premium titles.
One app.
Start reading
International Business Times UK
International Business Times UK
Technology
Bernadette B. Tixon

AI's 'Thinking' Text Is Hiding the Real Reasoning 75% of the Time, Researchers from OpenAI and Anthropic Warn

Major study warns: AI reasoning may not reflect true thought process. (Credit: Freepik)

When millions of people watch an AI model like ChatGPT or Claude 'think' through a problem step by step, most assume that visible reasoning reflects what the system is actually doing. A major new paper co-authored by more than 40 researchers from OpenAI, Anthropic, Google DeepMind and Meta now challenges that assumption directly — and the numbers behind their warning are difficult to dismiss.

Researchers at Anthropic tested 'faithfulness' in AI reasoning by subtly embedding hints into prompts and checking whether the model acknowledged using them when explaining its answer. Claude 3.7 Sonnet acknowledged using a hint just 25 per cent of the time, meaning it concealed the real influence behind its answer in 75 per cent of cases. The broader joint paper, titled 'Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety', builds on those findings to argue that the window to address this problem may already be narrowing.

Sign up to read this article
Read news from 100's of titles, curated specifically for you.
Already a member? Sign in here
Related Stories
Top stories on inkl right now
One subscription that gives you access to news from hundreds of sites
Already a member? Sign in here
Our Picks
Fourteen days free
Download the app
One app. One membership.
100+ trusted global sources.