Gemini 2.0 Flash Exp
Gemini 2.0 Flash Exp is behind the leaders in multimodal tasks. Too few results yet to rate reasoning, coding, agentic tasks, safety, long context, math, multilingual tasks, instruction following, or factuality.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Reasoning, Coding, Agentic, Safety, Long Context, Math, Multilingual, Instruction Following or Factuality.
Price
No current price is tracked for this model. See the rate card
Evidence
53results on49benchmarks
- 11 independently verified
- 4 aggregator
- 38 cross-referenced
From 12 sources · latest Oct 8, 2026 · How verification works
Gemini 2.0 Flash Exp benchmark results
53 results on 49 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
39.3% behind the leader4 of 6 ranked benchmarks measured
- 846.00Jun 5, 2026
- 70.60Jun 5, 2026
- 73.10Jun 5, 2026
- 57.00Jun 5, 2026
0 of 6 ranked benchmarks measured
- 18.90May 15, 2026
- GPQA Diamond63.64Oct 8, 2026gpqa
- Humanity's Last Exam4.07Oct 8, 2026aa_hle
Show 1 more reasoning resultHide 1 reasoning result
- GPQA Diamond62.10Jun 6, 2026GPQA (diamond)
0 of 10 ranked benchmarks measured
- LiveBench · Coding53.13Aug 23, 2026livebench_coding@2025-04-07
- SciCode34.03Sep 4, 2026aa_scicode
0 of 3 ranked benchmarks measured
- LiveBench · Instruction Following80.50Aug 23, 2026livebench_instruction_following@2025-04-07
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- livebench_language43.04Aug 23, 2026livebench_language@2025-04-07
- 94.60May 10, 2026
- 44.31May 10, 2026
- 78.30May 10, 2026
- 71.66May 10, 2026
- 90.13May 10, 2026
Show 36 more resultsHide 36 results
- AA Intelligence8.22Oct 8, 2026aa_intelligence_index
- 85.10Jun 5, 2026
- 22.20May 1, 2026
- 72.70Jun 6, 2026
- 88.30Jun 5, 2026
- 88.30Jun 12, 2026
- Chinese SimpleQA (C-SimpleQA)63.30Jun 6, 2026C-SimpleQA
- 92.90Jun 5, 2026
- 89.30Jun 6, 2026
- 94.60May 10, 2026
- 95.40Jun 6, 2026
- 89.60Jun 6, 2026
- 88.40Jun 12, 2026
- 88.40Jun 6, 2026
- M-LongDoc (multimodal long-document benchmark)31.40Jun 12, 2026M-LongDoc_acc
- 83.90Jun 6, 2026
- 75.90Jun 6, 2026
- 53.90Jun 12, 2026
- 86.50Jun 6, 2026
- 76.40Jun 6, 2026
- 53.30Jun 6, 2026
- 49.50Jun 6, 2026
- 12.20Jun 6, 2026
- 57.00Jun 6, 2026
- 57.50Jun 6, 2026
- 33.80Jun 6, 2026
- OlympiadBench46.10Jun 12, 2026OlympiadBench_full
- 96.00Jun 6, 2026
- 96.00Jun 6, 2026
- 95.10Jun 6, 2026
- 95.70Jun 6, 2026
- 93.70Jun 6, 2026
- 86.00Jun 6, 2026
- 79.70Jun 6, 2026
- 70.90Jun 6, 2026
- 26.60Jun 6, 2026
Gemini 2.0 Flash Exp: common questions
Who makes Gemini 2.0 Flash Exp?
Gemini 2.0 Flash Exp is made by Google.
When was Gemini 2.0 Flash Exp released?
Gemini 2.0 Flash Exp was released on Dec 11, 2024, according to Artificial Analysis.
What is Gemini 2.0 Flash Exp good at?
Gemini 2.0 Flash Exp is behind the leaders in multimodal tasks. Too few results yet to rate reasoning, coding, agentic tasks, safety, long context, math, multilingual tasks, instruction following, or factuality.
How many benchmarks has Gemini 2.0 Flash Exp been tested on?
We track 53 results for Gemini 2.0 Flash Exp on 49 benchmarks from 12 sources, 11 of them independently verified. The latest was recorded on Oct 8, 2026.
About this record
Where Gemini 2.0 Flash Exp's numbers come from, and every name it appears under.
- Tracked since
- May 1, 2026
- Newest source mention
- Sep 10, 2026
Where the results come from
Verification: 53 scores · 11 independently verified · 4 aggregator-attributed · 38 vendor cross-reference. How these tiers are assigned
From 12 sources on 5 sites. Hugging Face supplies 41 of them; the 11 independently verified results come from 4 sites. Bars are coloured by trust tier.
- huggingface.co41
- storage.googleapis.com6
- artificialanalysis.ai4
- aider.chat1
- simple-bench.com1