
With the recent launch of Claude 4.5 I've been testing it a lot. I recently put Claude 4.5 to the test against ChatGPT-5 and couldn’t believe the results. Anthropic calls their latest model the “smartest model yet,” which is why I couldn’t wait to see what it could do against Google’s Gemini 2.5 Pro.
To find how how the two compare, I put them through nine different challenges designed to stress-test their accuracy, reasoning and creativity — something these models are known for doing well in benchmark tests.
From logic and math word problems to coding and creative writing, here’s what I discovered when these two cutting-edge models went toe to toe. The results might surprise you!