Gemini 3 Pro
Gemini 3 Pro is at the frontier in multimodal tasks; strong in long context; capable in reasoning, instruction following, and factuality; and behind the leaders in coding and agentic tasks. Too few results yet to rate safety, math, or multilingual tasks.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Safety, Math or Multilingual.
Price
$2.00input$12.00outputper million tokens
From Artificial Analysis · All prices
Evidence
194results on129benchmarks
- 41 independently verified
- 31 aggregator
- 28 vendor-reported
- 94 cross-referenced
From 40 sources · latest Oct 8, 2026 · How verification works
Research
36 papers reference Gemini 3 ProGemini 3 Pro benchmark results
194 results on 129 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
Leads the field5 of 6 ranked benchmarks measured
- 88.40Jun 15, 2026
- MathVista89.80Jun 15, 2026MathVista (mini)
- 90.30Jun 15, 2026
- MMMU-Pro80.17Oct 8, 2026aa_mmmu_pro
- CharXiv (reasoning)81.40Oct 7, 2026CharXiv-R
Show 6 more multimodal resultsHide 6 multimodal results
- CharXiv (reasoning)81.40Jun 15, 2026CharXiv (RQ)
- 1304.91Aug 25, 2026
- 1290Jun 17, 2026
- 81.00Oct 7, 2026
- 81.00Jun 15, 2026
- 63.40May 22, 2026
8.6% behind the leader3 of 3 ranked benchmarks measured
- 65.60Jun 13, 2026
- 76.00Oct 8, 2026
- 77.00May 1, 2026
Show 2 more long context resultsHide 2 long context results
- 74.00Oct 8, 2026
- 68.20Jun 15, 2026
18.0% behind the leader5 of 6 ranked benchmarks measured
- GPQA Diamond90.81Oct 8, 2026gpqa
- 76.40May 15, 2026
- Humanity's Last Exam39.71Oct 8, 2026aa_hle
- ARC-AGI-231.10Oct 7, 2026ARC-AGI v2
- 9.14Oct 8, 2026
Show 10 more reasoning resultsHide 10 reasoning results
- 31.11May 10, 2026
- 0.00Oct 8, 2026
- GPQA Diamond88.69Oct 8, 2026gpqa
- GPQA Diamond91.90Oct 7, 2026GPQA
- GPQA Diamond91.00Aug 24, 2026GPQA-D
- 91.90Jul 29, 2026
- Humanity's Last Exam29.47Oct 8, 2026aa_hle
- 45.80Oct 7, 2026
- Humanity's Last Exam37.20Aug 24, 2026HLE w/o tools
- Humanity's Last Exam37.50Jun 15, 2026HLE-Full
21.6% behind the leader2 of 3 ranked benchmarks measured
- IFBench70.41Oct 8, 2026aa_ifbench
- 65.67Oct 8, 2026
23.2% behind the leader4 of 4 ranked benchmarks measured
- 72.90May 20, 2026
- 13.60May 2, 2026
- AA-Omniscience · Accuracy55.75Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination8.55Oct 8, 2026omniscienceNonHallucination
Show 2 more factuality resultsHide 2 factuality results
- AA-Omniscience · Accuracy48.32Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination10.00Oct 8, 2026omniscienceNonHallucination
26.4% behind the leader7 of 10 ranked benchmarks measured
- 90.70Jun 13, 2026
- SciCode56.13Sep 4, 2026aa_scicode
- 76.20Oct 7, 2026
- 65.00Aug 24, 2026
- Terminal-Bench Hard41.67Oct 8, 2026aa_terminalbench_hard
- 1439.58May 22, 2026
- 43.30Oct 8, 2026
Show 15 more coding resultsHide 15 coding results
- LiveCodeBench v687.40Jun 15, 2026LiveCodeBench (v6)
- SciCode49.88Sep 4, 2026aa_scicode
- 56.00May 1, 2026
- 56.00Aug 24, 2026
- 56.10Jun 15, 2026
- 68.70May 1, 2026
- 43.30May 1, 2026
- 69.60Sep 25, 2026
- 74.20May 1, 2026
- 76.20Jul 29, 2026
- SWE-bench Verified71.80Jun 5, 2026SWE-bench Verified (mini-swe-agent)
- 78.00Jun 5, 2026
- Terminal-Bench Hard34.09Oct 8, 2026aa_terminalbench_hard
- 56.90May 1, 2026
- 39.00Jun 13, 2026
38.6% behind the leader3 of 7 ranked benchmarks measured
- 70.30Oct 8, 2026
- 37.80Jun 15, 2026
- 34.21Jun 15, 2026
Show 5 more agentic resultsHide 5 agentic results
- 18.40May 1, 2026
- 59.20May 1, 2026
- 33.33Jun 15, 2026
- 54.10May 1, 2026
- MCP Atlas66.60May 17, 2026MCP-Atlas (Public Set)
0 of 5 ranked benchmarks measured
- 86.36Sep 2, 2026
- 91.67Sep 2, 2026
- 86.36May 17, 2026
Show 3 more math resultsHide 3 math results
- 91.67May 17, 2026
- 83.10Jun 15, 2026
- 83.30May 18, 2026
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- 1485Oct 8, 2026
- 63.80Sep 11, 2026
- 97.33Sep 7, 2026
- 97.49Sep 7, 2026
- 73.21Sep 7, 2026
- 72.50Sep 7, 2026
Show 111 more resultsHide 111 results
- 51.98Jun 18, 2026
- 45.05Jun 18, 2026
- AA Intelligence27.96Oct 8, 2026aa_intelligence_index
- AA Intelligence22.32Oct 8, 2026aa_intelligence_index
- 15.28Oct 8, 2026
- 1.80Oct 8, 2026
- 95.00Jul 5, 2026
- 100.00Oct 7, 2026
- 95.00Jun 15, 2026
- AIME 202596.00Jun 5, 2026AIME25
- 90.60May 17, 2026
- AIME25 no tools96.00Aug 24, 2026AIME25
- 97.05Sep 7, 2026
- 75.00May 10, 2026
- 31.10Jul 29, 2026
- 93.60Jun 13, 2026
- 72.60Jun 13, 2026
- 39.36Jun 18, 2026
- 46.49Jun 18, 2026
- 98.40Sep 7, 2026
- BrowseComp (context management)67.60Aug 24, 2026BrowseComp (w/ctx manage)
- 59.20Jun 5, 2026
- browsecomp_with_context_manager59.20Jun 13, 2026BrowseComp (w/ Context Manage)
- 66.80May 17, 2026
- 39.90Aug 24, 2026
- DeepSearchQA (F1)63.20Jun 15, 2026DeepSearchQA
- 49.90Jun 15, 2026
- frontiermath_tier_4_v118.75May 20, 2026frontiermath_tier_4
- 1195.00May 1, 2026
- HLE (with tools)45.80Jun 15, 2026HLE-Full (w/ tools)
- 97.50May 17, 2026
- HMMT Feb. 202597.30Jun 15, 2026HMMT 2025 (Feb)
- 97.50Jun 13, 2026
- 93.33May 17, 2026
- 93.30May 18, 2026
- 93.00May 17, 2026
- 57.20Jun 15, 2026
- 73.39Oct 7, 2026
- 92.00Jun 5, 2026
- 2439.00May 1, 2026
- 77.70Jun 15, 2026
- 73.50Jun 15, 2026
- 86.10Jun 15, 2026
- MathArena Apex (Pass@1)23.40Oct 7, 2026MathArena Apex
- 84.20Sep 2, 2026
- 71.34May 10, 2026
- 35.00May 10, 2026
- 99.49May 10, 2026
- 80.68Jun 15, 2026
- 90.10Jun 15, 2026
- 90.00Jun 5, 2026
- 91.80Oct 7, 2026
- 91.80Jul 29, 2026
- 78.88Jun 15, 2026
- 77.50Jun 15, 2026
- 89.39Jun 15, 2026
- 70.30Jun 15, 2026
- 89.70Jun 13, 2026
- 26.30May 1, 2026
- 42.70Jun 5, 2026
- 22.90Jun 5, 2026
- 68.50Jun 15, 2026
- 88.50Jun 15, 2026
- OmniDocBench v1.5 FormulaCDM89.18Sep 2, 2026
- OmniDocBench v1.5 Overall Score90.33Sep 2, 2026
- OmniDocBench v1.5 R-orderEdit0.07Sep 2, 2026
- OmniDocBench v1.5 TableTEDs88.28Sep 2, 2026
- OmniDocBench v1.5 TableTEDss90.29Sep 2, 2026
- OmniDocBench v1.5 TextEdit0.07Sep 2, 2026
- 75.83Sep 2, 2026
- 81.46Jun 15, 2026
- 82.85Jun 15, 2026
- ScreenSpot-Pro (No tools)72.70Oct 7, 2026ScreenSpot Pro
- 45.50Jun 15, 2026
- 72.10Oct 7, 2026
- 69.70Jun 15, 2026
- 69.60May 1, 2026
- 74.20May 1, 2026
- 6.50Jun 5, 2026
- 79.70Jun 5, 2026
- 85.40Oct 7, 2026
- 54.20Oct 7, 2026
- 54.20Jul 29, 2026
- 36.40May 17, 2026
- 36.40Jun 5, 2026
- vectara_answer_rate99.40May 2, 2026Answer Rate
- vectara_avg_summary_length101.90May 2, 2026Average Summary Length (Words)
- vectara_factual_consistency86.40May 2, 2026Factual Consistency Rate
- 547816.00Oct 7, 2026
- 82.40Jun 5, 2026
- 78.70Jun 5, 2026
- 78.70Jun 5, 2026
- 75.80Jun 5, 2026
- 89.20Jun 5, 2026
- 89.50Jun 5, 2026
- 66.78Jun 15, 2026
- 87.60Oct 7, 2026
- 87.60Jun 15, 2026
- 89.78Jun 15, 2026
- 57.00Jun 15, 2026
- 47.40Jun 15, 2026
- 8.00Jun 15, 2026
- 12.00Jun 15, 2026
- 85.30May 6, 2026
- τ²-Bench98.00Jul 29, 2026τ²-Bench (Telecom)
- 90.70Jun 13, 2026
- τ²-Bench87.00Jun 5, 2026𝜏²-Bench Telecom
- 85.30Jul 29, 2026
- τ²-Bench Telecom (AA run)87.13Oct 8, 2026aa_tau2
- τ²-Bench Telecom (AA run)68.13Oct 8, 2026aa_tau2
- 98.00May 1, 2026
Gemini 3 Pro: common questions
Who makes Gemini 3 Pro?
Gemini 3 Pro is made by Google.
When was Gemini 3 Pro released?
Gemini 3 Pro was released on Nov 18, 2025, according to Artificial Analysis.
What is Gemini 3 Pro good at?
Gemini 3 Pro is at the frontier in multimodal tasks; strong in long context; capable in reasoning, instruction following, and factuality; and behind the leaders in coding and agentic tasks. Too few results yet to rate safety, math, or multilingual tasks.
How much does Gemini 3 Pro cost?
Gemini 3 Pro costs $2.00 per million input tokens and $12.00 per million output tokens, according to Artificial Analysis. At a mix of three input tokens to one output token, it costs more than 85% of the 330 priced models we track.
How many benchmarks has Gemini 3 Pro been tested on?
We track 194 results for Gemini 3 Pro on 129 benchmarks from 40 sources, 41 of them independently verified. The latest was recorded on Oct 8, 2026.
About this record
Where Gemini 3 Pro's numbers come from, and every name it appears under.
- Tracked since
- Apr 25, 2026
- Newest source mention
- Aug 29, 2026
Where the results come from
Verification: 194 scores · 41 independently verified · 31 aggregator-attributed · 94 vendor cross-reference · 28 vendor-reported. How these tiers are assigned
From 40 sources on 16 sites. Hugging Face supplies 87 of them; the 41 independently verified results come from 11 sites. Bars are coloured by trust tier.
- huggingface.co87
- artificialanalysis.ai31
- api.llm-stats.com16
- deepmind.google12
- matharena.ai9
- raw.githubusercontent.com7
- www-cdn.anthropic.com7
- storage.googleapis.com6
- swebench.com5
- labs.scale.com3
- 99franklin.github.io2
- arcprize.org2
- datasets-server.huggingface.co2
- epoch.ai2
- lmarena.ai2
- simple-bench.com1
Also known as
How our sources name Gemini 3 Pro at each reasoning setting.
| Setting | Short form | API id |
|---|---|---|
| low | gemini 3 pro preview (low) | gemini-3-pro-low |
| high | Gemini 3 Pro Preview (high) Gemini 3 Pro Thinking (High) | — |