Anthropic, a leading AI company, tells Axios that its most advanced systems are learning not just to reason like humans — but also to reflect on, and express, how they actually think.
- They're starting to be introspective, like humans, Anthropic researcher Jack Lindsey, who studies models' "brains," tells us.
Why it matters: These introspective capabilities could make the models safer — or, possibly, just better at pretending to be safe.