Get all your news in one place.
100's of premium titles.
One app.
Start reading

Anthropic's models show signs of introspection

Anthropic, a leading AI company, tells Axios that its most advanced systems are learning not just to reason like humans — but also to reflect on, and express, how they actually think.

  • They're starting to be introspective, like humans, Anthropic researcher Jack Lindsey, who studies models' "brains," tells us.

Why it matters: These introspective capabilities could make the models safer — or, possibly, just better at pretending to be safe.

Sign up to read this article
Read news from 100's of titles, curated specifically for you.
Already a member? Sign in here
Related Stories
Top stories on inkl right now
One subscription that gives you access to news from hundreds of sites
Already a member? Sign in here
Our Picks
Fourteen days free
Download the app
One app. One membership.
100+ trusted global sources.