Models / Google

Gemini 3 Pro Deep Think

Gemini 3 · Released 2025-12-01

Google's extended-reasoning Gemini 3 Pro variant. ARC-AGI-2 semi-private 45.1%. For hardest puzzle-style reasoning, not default chat.

Reasoning-heavy Gemini route when standard Pro isn't enough on ARC-style tasks.

Composite35
Context1M
Input / 1M
Output / 1M
Knowledge cutoff2025-01-31
Statuslightweight
Not enough families covered yet.Add more benchmark rows for Gemini 3 Pro Deep Think to see the family profile.
Benchmark placements

Where Gemini 3 Pro Deep Think places on each public benchmark source that publishes a row. Rank counts every model with a latest row in the same test — not a universal quality score.

BenchmarkFamilyRankScoreSourceDate
CritPtReasoning#8/ 1226Artificial Analysis — CritPt2026-09-01
ARC-AGIReasoning#11/ 1445ARC Prize leaderboard2026-05-15
Family context

The three highest-scoring models with pages in each capability family. Where Gemini 3 Pro Deep Think shows up, it's highlighted.

Newest receipts

Benchmark rows added to the public ledger for Gemini 3 Pro Deep Think in the last 120 days. Older rows live in the full table below.

2 recent
BenchmarkFamilyScoreSourceDays ago
CritPtReasoning25.7%Artificial Analysis — CritPt5
ARC-AGIReasoning45.1%ARC Prize leaderboard114
Family coverage

A high composite that hides a weak family is a trap. These bars surface the families where this model hasn't been publicly tested, and where it leads.

Reasoning45
Price vs. performance
0255075100$0.10$0.30$1$3$10$30input USD / 1M tokens · log scale →← intelligence

Hover any dot for name, score, and input price. Keyboard: tab through the top twelve, or use the ranking below.

AnthropicOpenAIMistralAlibabaDeepSeekGoogleMoonshotMetaMiniMaxxAIZhipuXiaomiOtherMeituanCohereNVIDIA

Full benchmark ledger

Every catalog benchmark for Gemini 3 Pro Deep Think. Scores link to the original source; gaps mean no public row exists yet.

2 sourced rows

2 of 37 catalog benchmarks have a sourced row for Gemini 3 Pro Deep Think.

Reasoning

Sources