Get all your news in one place.
100's of premium titles.
One app.
Start reading
Tom’s Guide
Tom’s Guide
Technology
Amanda Caswell

I gave Claude Opus 5 and Kimi K3 15 impossible prompts — the winner surprised me

Moonshot vs. Anthropic.

When you push frontier models past standard coding tests and into the reality of enterprise workflows, their true capabilities and quirks show. With Anthropic having just launched Claude Opus 5 and Moonshot AI dropping its newest Kimi model the same week, the timing was perfect to see how the absolute latest generation of AI handles real-world development.

I ran 15 grueling, highly technical prompts against these two brand-new heavyweight engines. From orchestrating multi-agent TDD loops to architecting secure browser extensions, the goal was to test their analytical limits right out of the gate. What I found was a fundamental split in how these models solve problems. One thinks like a strategic infrastructure architect; the other builds like a lead systems engineer. Here is how it all shook out.

Sign up to read this article
Read news from 100's of titles, curated specifically for you.
Already a member? Sign in here
Related Stories
Top stories on inkl right now
One subscription that gives you access to news from hundreds of sites
Already a member? Sign in here
Our Picks
Fourteen days free
Download the app
One app. One membership.
100+ trusted global sources.