Gemini 3.6 Flash
Gemini 3.6 Flash is strong in long context and factuality; capable in multimodal tasks, reasoning, and agentic tasks; and behind the leaders in instruction following, coding, and math. Too few results yet to rate safety or multilingual tasks.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Safety or Multilingual.
Price
$0.75input$3.75outputper million tokens
From Artificial Analysis · 3 providers tracked · All prices
Evidence
66results on55benchmarks
- 26 independently verified
- 16 aggregator
- 24 vendor-reported
From 16 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingStructured outputsReasoning
As listed by OpenRouter
Gemini 3.6 Flash benchmark results
66 results on 55 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
5.0% behind the leader2 of 3 ranked benchmarks measured
- MRCR v2 (8-needle, 128K)91.80Jul 22, 2026GDM-MRCR v2 (8-needle) (128k (average))
- 80.00Oct 8, 2026
6.9% behind the leader3 of 4 ranked benchmarks measured
- 66.20Aug 2, 2026
- AA-Omniscience · Accuracy49.97Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination44.37Oct 8, 2026omniscienceNonHallucination
15.1% behind the leader2 of 6 ranked benchmarks measured
- CharXiv (reasoning)89.40Oct 7, 2026CharXiv-R
- MMMU-Pro83.24Oct 8, 2026aa_mmmu_pro
Show 2 more multimodal resultsHide 2 multimodal results
- CharXiv (reasoning)85.20Jul 22, 2026CharXiv Reasoning (No tools)
- 1296.85Sep 21, 2026
17.7% behind the leader5 of 6 ranked benchmarks measured
- GPQA Diamond92.83Oct 8, 2026gpqa
- LiveBench · Reasoning85.15Oct 8, 2026livebench_reasoning@2026-06-25
- 60.42Sep 21, 2026
- Humanity's Last Exam40.82Oct 8, 2026aa_hle
- 10.57Oct 8, 2026
21.8% behind the leader4 of 7 ranked benchmarks measured
- 83.00Oct 7, 2026
- τ-Bench V3 · Banking29.90Oct 8, 2026tauBanking
- 39.30Oct 8, 2026
- Terminal-Bench 4.07.07Oct 8, 2026
30.8% behind the leader1 of 3 ranked benchmarks measured
- LiveBench · Instruction Following75.37Oct 8, 2026livebench_instruction_following@2026-06-25
31.8% behind the leader6 of 10 ranked benchmarks measured
- LiveBench · Coding77.86Oct 8, 2026livebench_coding@2026-06-25
- Terminal-Bench 2.177.53Oct 8, 2026terminalbenchV21
- SciCode53.36Oct 8, 2026aa_scicode
- 1537.76Sep 21, 2026
- 58.70Oct 7, 2026
- LiveBench · Agentic Coding43.43Oct 8, 2026livebench_agentic_coding@2026-06-25
Show 2 more coding resultsHide 2 coding results
- 1536.95Jul 22, 2026
- 78.00Oct 7, 2026
50.9% behind the leader5 of 5 ranked benchmarks measured
- LiveBench · Mathematics86.40Oct 8, 2026livebench_math@2026-06-25
- 58.95Sep 9, 2026
- 21.95Sep 8, 2026
Show 2 more math resultsHide 2 math results
- 96.67Sep 2, 2026
- 89.39Sep 2, 2026
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- livebench_language83.90Oct 8, 2026livebench_language@2026-06-25
- livebench_data_analysis63.00Oct 8, 2026livebench_data_analysis@2026-06-25
- 1483Oct 6, 2026
- 34.50Sep 22, 2026
- AA Intelligence34.00Sep 21, 2026Artificial Analysis Intelligence Index
- 76.50Sep 21, 2026
Show 25 more resultsHide 25 results
- 30.15Sep 9, 2026
- AA Intelligence33.98Oct 8, 2026aa_intelligence_index
- 22.13Oct 8, 2026
- 24.20Aug 14, 2026
- 83.17Sep 21, 2026
- 91.17Sep 21, 2026
- Artificial Analysis Coding Index69.24Sep 9, 2026aa_coding_index
- CharXiv Reasoning (no tools)89.40Aug 24, 2026CharXiv Reasoning with tools
- 89.40Jul 22, 2026
- 1538.00Aug 14, 2026
- 49.00Aug 4, 2026
- 49.00Oct 7, 2026
- DeepSWE 1.148.60Jul 22, 2026DeepSWE v1.1
- 54.00Jul 22, 2026
- 54.00Aug 24, 2026
- 91.80Aug 24, 2026
- GDP (Surge AI)22.00Aug 14, 2026GDP.pdf
- 1421.00Aug 24, 2026
- GDPval-AA v2 Elo1422.00Jul 22, 2026GDPVal-AA v2 (Elo)
- 51.20Aug 14, 2026
- 84.20Aug 14, 2026
- 87.85Sep 2, 2026
- 63.90Oct 7, 2026
- 33.80Aug 14, 2026
- 5.40Aug 14, 2026
Gemini 3.6 Flash: common questions
Who makes Gemini 3.6 Flash?
Gemini 3.6 Flash is made by Google.
When was Gemini 3.6 Flash released?
Gemini 3.6 Flash was released on Jul 21, 2026, according to Artificial Analysis.
What is Gemini 3.6 Flash good at?
Gemini 3.6 Flash is strong in long context and factuality; capable in multimodal tasks, reasoning, and agentic tasks; and behind the leaders in instruction following, coding, and math. Too few results yet to rate safety or multilingual tasks.
How much does Gemini 3.6 Flash cost?
Gemini 3.6 Flash costs $0.75 per million input tokens and $3.75 per million output tokens, according to Artificial Analysis. We track its price at 3 providers. At a mix of three input tokens to one output token, it costs more than 65% of the 330 priced models we track.
How many benchmarks has Gemini 3.6 Flash been tested on?
We track 66 results for Gemini 3.6 Flash on 55 benchmarks from 16 sources, 26 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does Gemini 3.6 Flash support?
OpenRouter lists tool calling, structured outputs, and reasoning for Gemini 3.6 Flash.
About this record
Where Gemini 3.6 Flash's numbers come from, and every name it appears under.
- Tracked since
- Jul 21, 2026
- Newest source mention
- Aug 23, 2026
Where the results come from
Verification: 66 scores · 26 independently verified · 16 aggregator-attributed · 24 vendor-reported. How these tiers are assigned
From 16 sources on 11 sites. Artificial Analysis supplies 17 of them; the 26 independently verified results come from 7 sites. Bars are coloured by trust tier.
- artificialanalysis.ai17
- deepmind.google13
- arcprize.org8
- livebench.ai7
- api.llm-stats.com6
- storage.googleapis.com4
- datasets-server.huggingface.co3
- epoch.ai3
- matharena.ai3
- blog.google1
- lmarena.ai1
Also known as
How our sources name Gemini 3.6 Flash at each reasoning setting.
| Setting | Short form | API id |
|---|---|---|
| minimal | gemini 3.6 flash (minimal) | — |
| low | gemini 3.6 flash (low) | — |
| medium | gemini 3.6 flash (medium) | — |
| high | gemini 3.6 flash (high) | gemini-3.6-flash-high |