Gemini 2.0 Flash (Reasoning)
Gemini 2.0 Flash (Reasoning) is behind the leaders in coding. Too few results yet to rate reasoning, agentic tasks, safety, long context, math, multimodal tasks, multilingual tasks, instruction following, or factuality.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Reasoning, Agentic, Safety, Long Context, Math, Multimodal, Multilingual, Instruction Following or Factuality.
Price
No current price is tracked for this model. See the rate card
Evidence
24results on23benchmarks
- 18 independently verified
- 6 vendor-reported
From 16 sources · latest Oct 7, 2026 · How verification works
Gemini 2.0 Flash (Reasoning) benchmark results
24 results on 23 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
45.2% behind the leader1 of 10 ranked benchmarks measured
- 13.52Sep 24, 2026
Show 1 more coding resultHide 1 coding result
- LiveBench · Coding53.91Aug 23, 2026livebench_coding@2025-04-07
0 of 6 ranked benchmarks measured
- 1.30Sep 20, 2026
- 30.70May 10, 2026
- GPQA Diamond74.20Oct 7, 2026GPQA
0 of 6 ranked benchmarks measured
- 1157.86Sep 13, 2026
- 75.40Oct 7, 2026
0 of 3 ranked benchmarks measured
- LiveBench · Instruction Following85.53Aug 23, 2026livebench_instruction_following@2025-04-07
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- 95.28Sep 20, 2026
- 98.49Sep 20, 2026
- 66.19Sep 20, 2026
- 95.40Sep 20, 2026
- 99.41Sep 20, 2026
- 66.22Sep 20, 2026
Show 10 more resultsHide 10 results
- 27.50Jul 5, 2026
- 71.50Oct 7, 2026
- 13.33May 11, 2026
- livebench_language45.17Aug 23, 2026livebench_language@2025-04-07
- 35.10Aug 23, 2026
- 76.40Oct 7, 2026
- 69.20Oct 7, 2026
- 13.52Aug 31, 2026
- 4.17Sep 2, 2026
- 0.00May 10, 2026
Gemini 2.0 Flash (Reasoning): common questions
Who makes Gemini 2.0 Flash (Reasoning)?
Gemini 2.0 Flash (Reasoning) is made by Google.
What is Gemini 2.0 Flash (Reasoning) good at?
Gemini 2.0 Flash (Reasoning) is behind the leaders in coding. Too few results yet to rate reasoning, agentic tasks, safety, long context, math, multimodal tasks, multilingual tasks, instruction following, or factuality.
How many benchmarks has Gemini 2.0 Flash (Reasoning) been tested on?
We track 24 results for Gemini 2.0 Flash (Reasoning) on 23 benchmarks from 16 sources, 18 of them independently verified. The latest was recorded on Oct 7, 2026.
About this record
Where Gemini 2.0 Flash (Reasoning)'s numbers come from, and every name it appears under.
- Tracked since
- Apr 27, 2026
- Newest source mention
- Jun 24, 2026
Where the results come from
Verification: 24 scores · 18 independently verified · 6 vendor-reported. How these tiers are assigned
From 16 sources on 8 sites. api.llm-stats.com supplies 6 of them; the 18 independently verified results come from 7 sites. Bars are coloured by trust tier.
- api.llm-stats.com6
- storage.googleapis.com6
- matharena.ai4
- huggingface.co3
- swebench.com2
- arcprize.org1
- datasets-server.huggingface.co1
- simple-bench.com1