Get all your news in one place.
100's of premium titles.
One app.
Start reading
Tom’s Guide
Tom’s Guide
Technology
Amanda Caswell

I tested ChatGPT-5.2 vs Claude 4.6 Opus in 9 tough challenges — here’s the winner

Chatgpt and claude logos on phones.

As someone who spends every day testing the "holes" in AI logic, I’ve been eagerly waiting to see how the landscape shifts with the release of Claude 4.6 Opus. We are no longer in the era where "it works" is enough; we are looking for nuance, meta-awareness and the ability to handle the messy contradictions of human thought.

To see if Anthropic’s latest flagship lives up to the hype, I put it head-to-head against ChatGPT-5.2 Thinking in a nine-round "Reasoning Gauntlet." My goal wasn't just to find the right answers — it was to find the most "human" ones. I tested them on everything from counterintuitive physics and ethical trade-offs to the "show, don't tell" math problems that usually trip up LLMs. This wasn't just a benchmark; it was an attempt to see which model truly understands the why behind the what.

Sign up to read this article
Read news from 100's of titles, curated specifically for you.
Already a member? Sign in here
Related Stories
Top stories on inkl right now
One subscription that gives you access to news from hundreds of sites
Already a member? Sign in here
Our Picks
Fourteen days free
Download the app
One app. One membership.
100+ trusted global sources.