Google Launches Gemini 2.5 Pro Deep Think — New Benchmark Leader
Google's Gemini 2.5 Pro with Deep Think mode tops science and reasoning benchmarks, using parallel thinking to outperform GPT-5.5 and Fable 5.
Quick answer
Google launched Gemini 2.5 Pro with Deep Think on June 22, 2026. The model scores 82.4% on GPQA Diamond and 89.8% on MMLU-Pro — the highest of any publicly available model — by exploring multiple reasoning paths in parallel before answering. It is available now in the Gemini app.
Google Launches Gemini 2.5 Pro Deep Think — New Benchmark Leader
Google has launched Gemini 2.5 Pro with Deep Think, a new reasoning mode that makes it the top-performing AI model on science and graduate-level reasoning benchmarks. The release, which rolled out on June 22, represents the most capable publicly available model Google has ever shipped.
What Deep Think Does Differently
Most AI models generate a single chain of reasoning and commit to it. Deep Think takes a different approach: it explores multiple hypotheses in parallel, revises and combines them, then settles on the strongest answer. Google calls this “parallel thinking,” and it is visible to users — you can see thought summaries showing which paths the model considered and why it chose its final response.
The result is a model that excels at problems requiring careful deliberation. On GPQA Diamond, a benchmark of graduate-level physics, chemistry, and biology questions, Gemini 2.5 Pro Deep Think scored 82.4% — ahead of Anthropic’s Fable 5 at 79.1% and OpenAI’s GPT-5.5 at 76.3%. On MMLU-Pro, a broad knowledge benchmark, it scored 89.8%, the highest of any publicly available model.
Where It Still Trails
The model is not the best at everything. Anthropic’s Fable 5 still leads on software engineering benchmarks, scoring 88.6% on SWE-bench Verified compared to Gemini’s lower marks in that category. GPT-5.5 maintains an edge in creative writing and conversational naturalness. Each frontier model now has a clear speciality rather than dominating across the board.
What This Means for You
If you use AI for research, study, or technical problem-solving, Deep Think is worth trying. It is particularly strong at breaking down complex scientific questions, working through multi-step maths problems, and analysing code.
To use it, open the Gemini app, select Gemini 2.5 Pro from the model picker, and toggle “Deep Think” on in the prompt bar. It is available to Gemini Advanced subscribers.
This launch arrives during the most competitive stretch in AI history. With GPT-5.6 expected any day and Anthropic rapidly expanding its research team, the gap between leading models continues to narrow. For everyday users, the real winner is the growing range of capable tools to choose from — whether you need deep reasoning from Gemini, coding help from Claude, or creative writing from ChatGPT.
Sign up for our newsletter to get daily AI news delivered to your inbox.
Frequently asked questions
What is Gemini 2.5 Pro Deep Think?
How do I use Gemini Deep Think mode?
How does Gemini 2.5 Pro Deep Think compare to GPT-5.5 and Claude?
Is Gemini 2.5 Pro Deep Think free to use?
What is the context window for Gemini 2.5 Pro?
Want to keep learning?
Explore our guided learning paths or try building something with AI right now.
More from News
Claude Cowork Expands to Mobile and Web
Claude Cowork Expands to Mobile and Web
Anthropic brings Claude Cowork to web and mobile, letting AI sessions run remotely and continue across devices without a laptop.
OpenAI Launches ChatGPT Work — an AI Agent for Your Job
OpenAI Launches ChatGPT Work — an AI Agent for Your Job
OpenAI debuts ChatGPT Work, an autonomous agent that completes tasks across your apps and delivers finished documents, spreadsheets, and slides.
SpaceXAI and Cursor Launch Grok 4.5 — A Coding-First AI Model
SpaceXAI and Cursor Launch Grok 4.5 — A Coding-First AI Model
SpaceXAI and Cursor publicly launch Grok 4.5, a joint AI model built for coding, legal, and finance tasks. Available today.
Enjoyed this article?
Subscribe for more AI insights delivered to your inbox every week.