Gemini 2.5 Pro
Gemini 2.5 Pro is capable in multimodal tasks and factuality; and behind the leaders in long context, reasoning, coding, instruction following, and math. Too few results yet to rate agentic tasks, safety, or multilingual tasks.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Agentic, Safety or Multilingual.
Price
$1.25input$10.00outputper million tokens
From Artificial Analysis · 4 providers tracked · All prices
Evidence
163results on92benchmarks
- 66 independently verified
- 28 aggregator
- 24 vendor-reported
- 45 cross-referenced
From 38 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingStructured outputsReasoning
As listed by OpenRouter
Gemini 2.5 Pro benchmark results
163 results on 92 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
19.0% behind the leader5 of 6 ranked benchmarks measured
- 84.80Oct 7, 2026
- 85.90Jun 5, 2026
- 79.60Oct 7, 2026
- 83.88Jun 5, 2026
- MMMU-Pro74.91Oct 8, 2026aa_mmmu_pro
Show 6 more multimodal resultsHide 6 multimodal results
- 1263.23Aug 25, 2026
- 1246Jun 17, 2026
- 82.00Oct 7, 2026
- 83.89Jun 5, 2026
- 79.18Jun 5, 2026
- 59.30May 22, 2026
21.9% behind the leader4 of 4 ranked benchmarks measured
- 7.00May 2, 2026
- 56.00May 20, 2026
- AA-Omniscience · Accuracy39.05Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination9.11Oct 8, 2026omniscienceNonHallucination
26.2% behind the leader2 of 3 ranked benchmarks measured
- 58.80Jun 5, 2026
- 69.00Oct 8, 2026
Show 1 more long context resultHide 1 long context result
- 65.00Jun 5, 2026
34.8% behind the leader5 of 6 ranked benchmarks measured
- GPQA Diamond83.64Oct 8, 2026gpqa
- 51.60May 10, 2026
- Humanity's Last Exam18.03Oct 8, 2026aa_hle
- 2.57Oct 8, 2026
- ARC-AGI-24.90Oct 7, 2026ARC-AGI v2
Show 18 more reasoning resultsHide 18 reasoning results
- 4.86Sep 22, 2026
- 4.03Sep 22, 2026
- 2.92Sep 22, 2026
- 0.00Sep 22, 2026
- 3.75May 10, 2026
- GPQA Diamond84.44Oct 8, 2026gpqa
- GPQA Diamond82.22Oct 8, 2026gpqa
- GPQA Diamond86.40Oct 7, 2026GPQA
- GPQA Diamond83.00Oct 7, 2026GPQA
- 84.00Jun 6, 2026
- 86.40Jun 5, 2026
- Humanity's Last Exam22.52Oct 8, 2026aa_hle
- Humanity's Last Exam19.93Oct 8, 2026aa_hle
- 21.60Oct 7, 2026
- 17.80Oct 7, 2026
- Humanity's Last Exam21.10Jun 6, 2026HLE (w/o tools)
- Humanity's Last Exam21.60Jun 5, 2026HLE (no tools)
- 62.40May 10, 2026
38.3% behind the leader5 of 10 ranked benchmarks measured
- SciCode46.30Oct 8, 2026aa_scicode
- 63.20Oct 7, 2026
- Terminal-Bench Hard26.52Oct 8, 2026aa_terminalbench_hard
- 1224.23May 22, 2026
- Terminal-Bench 2.128.46Oct 8, 2026terminalbenchV21
Show 10 more coding resultsHide 10 coding results
- LiveBench · Coding86.72Jun 17, 2026livebench_coding@2025-04-07
- SciCode39.47Sep 4, 2026aa_scicode
- SciCode41.55Sep 4, 2026aa_scicode
- 43.00Jun 6, 2026
- 53.60Sep 1, 2026
- 67.20Oct 7, 2026
- 67.20May 3, 2026
- 63.80Jun 6, 2026
- 67.20Jun 5, 2026
- 25.00Jun 6, 2026
40.9% behind the leader2 of 3 ranked benchmarks measured
- 53.62Oct 8, 2026
- IFBench48.71Oct 8, 2026aa_ifbench
Show 3 more instruction following resultsHide 3 instruction following results
- 49.00Jun 6, 2026
- LiveBench · Instruction Following80.95Jun 17, 2026livebench_instruction_following@2025-04-07
- 51.80Jun 5, 2026
56.2% behind the leader2 of 5 ranked benchmarks measured
- 24.56Sep 9, 2026
- 0.00Sep 8, 2026
0 of 7 ranked benchmarks measured
- 0.00Oct 8, 2026
- Terminal-Bench 4.00.00Oct 8, 2026
- τ-Bench V3 · Banking9.69Oct 8, 2026tauBanking
Show 1 more agentic resultHide 1 agentic result
- 9.90Jun 6, 2026
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- 37.00Sep 22, 2026
- 41.00Sep 22, 2026
- 29.50Sep 22, 2026
- 16.00Sep 22, 2026
- 62.20Sep 11, 2026
- 79.10Sep 3, 2026
Show 90 more resultsHide 90 results
- 3.55Sep 9, 2026
- AA Intelligence16.00Jun 19, 2026Artificial Analysis Intelligence Index
- AA Intelligence26.00Jun 18, 2026Artificial Analysis Intelligence Index
- AA Intelligence14.96Oct 8, 2026aa_intelligence_index
- AA Intelligence14.52Oct 8, 2026aa_intelligence_index
- AA Intelligence16.08Oct 8, 2026aa_intelligence_index
- -16.35Oct 8, 2026
- 72.90May 1, 2026
- 76.90May 1, 2026
- 82.20Oct 7, 2026
- 76.50Oct 7, 2026
- 82.20May 3, 2026
- 92.00Jun 5, 2026
- 83.33Jul 5, 2026
- 88.33Jul 5, 2026
- 88.00Oct 7, 2026
- 83.00Oct 7, 2026
- AIME 202588.00Jun 6, 2026AIME25
- 83.96Jun 5, 2026
- 88.00Jun 5, 2026
- 73.59May 19, 2026
- 99.50May 10, 2026
- 33.00May 10, 2026
- 1446Jul 23, 2026
- 57.70Jun 6, 2026
- Artificial Analysis Coding Index46.73Sep 9, 2026aa_coding_index
- Artificial Analysis Coding Index33.25Sep 9, 2026aa_coding_index
- 96.40May 10, 2026
- 32.20Jun 6, 2026
- 42.60Jun 6, 2026
- frontiermath_tier_4_v14.17May 20, 2026frontiermath_tier_4
- 60.20Jun 6, 2026
- 89.20Oct 7, 2026
- 88.60Oct 7, 2026
- 65.44May 10, 2026
- HLE (with tools)28.40Jun 6, 2026HLE (w/ tools)
- 65.68Jun 5, 2026
- 82.50May 17, 2026
- 80.83May 11, 2026
- 80.00May 17, 2026
- 88.00Sep 2, 2026
- 31.55Sep 2, 2026
- 8.33May 17, 2026
- livebench_language65.93Jun 17, 2026livebench_language@2025-04-07
- 84.27May 30, 2026
- 82.75May 3, 2026
- 82.75May 3, 2026
- 74.20May 3, 2026
- LiveCodeBench80.00Jun 6, 2026LiveCodeBench (LCB)
- 72.01Jun 5, 2026
- 77.10Jun 5, 2026
- LiveCodeBench (v5)75.60Oct 7, 2026LiveCodeBench v5
- 99.07May 30, 2026
- 98.76May 3, 2026
- 98.76May 3, 2026
- 62.00May 30, 2026
- 59.43May 3, 2026
- 59.43May 3, 2026
- 92.17May 30, 2026
- 90.60May 3, 2026
- 90.60May 3, 2026
- MATH-500 (EM)98.80Jun 5, 2026MATH-500
- 73.30Jun 5, 2026
- 77.22Sep 2, 2026
- 93.19Jun 5, 2026
- 86.00Jun 6, 2026
- 86.00Jun 5, 2026
- 93.00Oct 7, 2026
- 58.80Jun 5, 2026
- 76.80Jun 5, 2026
- 96.98May 10, 2026
- 54.00Oct 7, 2026
- 50.80Oct 7, 2026
- 54.00Jun 5, 2026
- 53.60May 1, 2026
- 50.00Jun 5, 2026
- 67.00Jun 5, 2026
- 25.30Jun 6, 2026
- theagentcompany39.30Jun 6, 2026AgentCompany
- 24.40Sep 2, 2026
- 16.67May 17, 2026
- vectara_answer_rate99.10May 2, 2026Answer Rate
- vectara_avg_summary_length106.40May 2, 2026Average Summary Length (Words)
- vectara_factual_consistency93.00May 2, 2026Factual Consistency Rate
- 83.60Oct 7, 2026
- 56.00Jun 6, 2026
- 98.67May 10, 2026
- 91.60Jun 5, 2026
- τ²-Bench54.00Jun 6, 2026τ²-Bench-Telecom
- τ²-Bench Telecom (AA run)54.09Oct 8, 2026aa_tau2
Gemini 2.5 Pro: common questions
Who makes Gemini 2.5 Pro?
Gemini 2.5 Pro is made by Google.
When was Gemini 2.5 Pro released?
Gemini 2.5 Pro was released on Jun 5, 2025, according to Artificial Analysis.
What is Gemini 2.5 Pro good at?
Gemini 2.5 Pro is capable in multimodal tasks and factuality; and behind the leaders in long context, reasoning, coding, instruction following, and math. Too few results yet to rate agentic tasks, safety, or multilingual tasks.
How much does Gemini 2.5 Pro cost?
Gemini 2.5 Pro costs $1.25 per million input tokens and $10.00 per million output tokens, according to Artificial Analysis. We track its price at 4 providers. At a mix of three input tokens to one output token, it costs more than 78% of the 330 priced models we track.
How many benchmarks has Gemini 2.5 Pro been tested on?
We track 163 results for Gemini 2.5 Pro on 92 benchmarks from 38 sources, 66 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does Gemini 2.5 Pro support?
OpenRouter lists tool calling, structured outputs, and reasoning for Gemini 2.5 Pro.
About this record
Where Gemini 2.5 Pro's numbers come from, and every name it appears under.
- Tracked since
- Apr 27, 2026
- Newest source mention
- Sep 10, 2026
Where the results come from
Verification: 163 scores · 66 independently verified · 28 aggregator-attributed · 45 vendor cross-reference · 24 vendor-reported. How these tiers are assigned
From 38 sources on 16 sites. Hugging Face supplies 48 of them; the 66 independently verified results come from 15 sites. Bars are coloured by trust tier.
- huggingface.co48
- artificialanalysis.ai30
- api.llm-stats.com21
- livecodebench.github.io12
- matharena.ai11
- arcprize.org10
- storage.googleapis.com9
- epoch.ai4
- raw.githubusercontent.com4
- aider.chat3
- 99franklin.github.io2
- datasets-server.huggingface.co2
- lmarena.ai2
- simple-bench.com2
- swebench.com2
- labs.scale.com1