Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

Gemini 3.5 Flash

Score basis

Gemini 3.5 Flash is strong in factuality; capable in instruction following, reasoning, multimodal tasks, and agentic tasks; and behind the leaders in coding, long context, and math. Too few results yet to rate safety or multilingual tasks.

Capability profile

Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.

Strong1 capability
  1. Factuality−7.8%3 of 4
Capable4 capabilities
  1. Instruction Following−11.1%2 of 3
  2. Reasoning−13.6%6 of 6
  3. Multimodal−15.3%3 of 6
  4. Agentic−18.5%6 of 7
Limited3 capabilities
  1. Coding−25.1%7 of 10
  2. Long Context−25.7%2 of 3
  3. Math−46.0%5 of 5
Not rated2 capabilities

Too few results yet to rate Safety or Multilingual.

Price

$1.50input$9.00outputper million tokens

From Artificial Analysis · 4 providers tracked · All prices

Costs more than 78% of 330 priced models · 3:1 input-to-output blend, log scale

Evidence

132results on75benchmarks

  • 29 independently verified
  • 53 aggregator
  • 29 vendor-reported
  • 21 cross-referenced

From 28 sources · latest Oct 8, 2026 · How verification works

API features

Tool callingStructured outputsReasoning

As listed by OpenRouter

Gemini 3.5 Flash benchmark results

132 results on 75 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.

Factuality

StrongFull factuality ranking

7.8% behind the leader3 of 4 ranked benchmarks measured

Show 4 more factuality resultsHide 4 factuality results

Instruction Following

CapableFull instruction following ranking

11.1% behind the leader2 of 3 ranked benchmarks measured

Show 2 more instruction following resultsHide 2 instruction following results

Reasoning

CapableFull reasoning ranking

13.6% behind the leader6 of 6 ranked benchmarks measured

Show 10 more reasoning resultsHide 10 reasoning results

Multimodal

CapableFull multimodal ranking

15.3% behind the leader3 of 6 ranked benchmarks measured

Show 5 more multimodal resultsHide 5 multimodal results

Agentic

CapableFull agentic ranking

18.5% behind the leader6 of 7 ranked benchmarks measured

Show 7 more agentic resultsHide 7 agentic results

Coding

LimitedFull coding ranking

25.1% behind the leader7 of 10 ranked benchmarks measured

Show 10 more coding resultsHide 10 coding results

Long Context

LimitedFull long context ranking

25.7% behind the leader2 of 3 ranked benchmarks measured

Show 3 more long context resultsHide 3 long context results

46.0% behind the leader5 of 5 ranked benchmarks measured

Show 3 more math resultsHide 3 math results

More results

Benchmarks outside the capability baskets. They are not ranked against a leader.

Show 51 more resultsHide 51 results

Gemini 3.5 Flash: common questions

Who makes Gemini 3.5 Flash?

Gemini 3.5 Flash is made by Google.

When was Gemini 3.5 Flash released?

Gemini 3.5 Flash was released on May 19, 2026, according to Artificial Analysis.

What is Gemini 3.5 Flash good at?

Gemini 3.5 Flash is strong in factuality; capable in instruction following, reasoning, multimodal tasks, and agentic tasks; and behind the leaders in coding, long context, and math. Too few results yet to rate safety or multilingual tasks.

How much does Gemini 3.5 Flash cost?

Gemini 3.5 Flash costs $1.50 per million input tokens and $9.00 per million output tokens, according to Artificial Analysis. We track its price at 4 providers. At a mix of three input tokens to one output token, it costs more than 78% of the 330 priced models we track.

How many benchmarks has Gemini 3.5 Flash been tested on?

We track 132 results for Gemini 3.5 Flash on 75 benchmarks from 28 sources, 29 of them independently verified. The latest was recorded on Oct 8, 2026.

Which API features does Gemini 3.5 Flash support?

OpenRouter lists tool calling, structured outputs, and reasoning for Gemini 3.5 Flash.

About this record

Where Gemini 3.5 Flash's numbers come from, and every name it appears under.

Tracked since
May 19, 2026
Newest source mention
Aug 24, 2026

Where the results come from

Verification: 132 scores · 29 independently verified · 53 aggregator-attributed · 21 vendor cross-reference · 29 vendor-reported. How these tiers are assigned

From 28 sources on 17 sites. Artificial Analysis supplies 54 of them; the 29 independently verified results come from 9 sites. Bars are coloured by trust tier.

  • artificialanalysis.ai54
  • api.llm-stats.com15
  • lf3-static.bytednsdoc.com10
  • www-cdn.anthropic.com8
  • deepmind.google7
  • livebench.ai7
  • datasets-server.huggingface.co5
  • storage.googleapis.com5
  • arcprize.org4
  • epoch.ai4
  • matharena.ai4
  • anthropic.com2
  • blog.google2
  • labs.scale.com2
  • cdn.sanity.io1
  • lmarena.ai1
  • simple-bench.com1

Also known as

How our sources name Gemini 3.5 Flash at each reasoning setting.

SettingShort formAPI id
minimalgemini 3.5 flash (minimal)gemini-3-5-flash-minimal
mediumgemini 3.5 flash (medium)gemini-3-5-flash-medium gemini-3.5-flash-medium
highgemini 3.5 flash (high)gemini-3.5-flash-high
Also listed asgemini flash 3.5gemini-3-5-flashgemini-3-5-flash (Non-Reasoning)gemini-3-5-flash (reasoning)gemini-3.5-flash:batch