Get all your news in one place.
100's of premium titles.
One app.
Start reading
Tom’s Guide
Tom’s Guide
Technology
Amanda Caswell

I tested Gemini 3 Flash vs Claude 4.6 Opus in 9 tough challenges — here’s the winner

Gemini vs claude .

Claude 4.6 Opus launched just days ago, and I immediately pitted it against ChatGPT-5.2 Thinking to see how it compared to OpenAI’s smartest model. Naturally, with Gemini’s recent dominance, I had to see how it compared to Gemini 3 Flash.

I put the two top models head-to-head across nine challenging tests spanning math, logic, coding, creative writing and more — tasks designed to push each model's reasoning, creativity and practical usefulness to the limit.

My prompts aren’t the kind of questions you can answer by regurgitating training data; they require genuine multi-step thinking, context judgment and the ability to follow complex constraints. Here's how Anthropic's most powerful model stacked up against Google's latest.

1. Multi-step math reasoning

Sign up to read this article
Read news from 100's of titles, curated specifically for you.
Already a member? Sign in here
Related Stories
Top stories on inkl right now
One subscription that gives you access to news from hundreds of sites
Already a member? Sign in here
Our Picks
Fourteen days free
Download the app
One app. One membership.
100+ trusted global sources.