Gemini Deep Think Tops The Benchmarks
22JUN
Google grabbed the top score while Anthropic's best models sit benched. Gemini 2.5 Deep Think beat GPT-5.5 and Fable 5 on graduate-level science. Timing is everything.
Deep Think uses parallel reasoning, running many thought paths at once. It scored 82.4% on GPQA Diamond, a hard science test. That beats GPT-5.5 and the suspended Fable 5.
The win lands while US export rules keep Anthropic's Fable 5 and Mythos offline. Google has a clear lane to claim the lead.
It is live for AI Ultra subscribers, with API access soon. Benchmarks are not products. But mindshare moves on leaderboard wins.