Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

Gemini 2.5 Flash (Sep) (Non-Reasoning)

Score basis

Gemini 2.5 Flash (Sep) (Non-Reasoning) is capable in long context; and behind the leaders in multimodal tasks, factuality, agentic tasks, instruction following, and reasoning. Too few results yet to rate coding, safety, math, or multilingual tasks.

Capability profile

Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.

Capable1 capability
  1. Long Context−24.4%1 of 3
Limited5 capabilities
  1. Multimodal−31.1%1 of 6
  2. Factuality−37.5%2 of 4
  3. Agentic−37.6%1 of 7
  4. Instruction Following−38.0%1 of 3
  5. Reasoning−40.7%3 of 6
Not rated4 capabilities

Too few results yet to rate Coding, Safety, Math or Multilingual.

Price

No current price is tracked for this model. See the rate card

Evidence

34results on17benchmarks

  • 2 independently verified
  • 32 aggregator

From 3 sources · latest Oct 8, 2026 · How verification works

Gemini 2.5 Flash (Sep) (Non-Reasoning) benchmark results

34 results on 17 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.

Long Context

CapableFull long context ranking

24.4% behind the leader1 of 3 ranked benchmarks measured

Show 1 more long context resultHide 1 long context result

Multimodal

LimitedFull multimodal ranking

31.1% behind the leader1 of 6 ranked benchmarks measured

Show 3 more multimodal resultsHide 3 multimodal results

Factuality

LimitedFull factuality ranking

37.5% behind the leader2 of 4 ranked benchmarks measured

Show 2 more factuality resultsHide 2 factuality results

Agentic

LimitedFull agentic ranking

37.6% behind the leader1 of 7 ranked benchmarks measured

Show 1 more agentic resultHide 1 agentic result

Instruction Following

LimitedFull instruction following ranking

38.0% behind the leader1 of 3 ranked benchmarks measured

Show 1 more instruction following resultHide 1 instruction following result

Reasoning

LimitedFull reasoning ranking

40.7% behind the leader3 of 6 ranked benchmarks measured

Show 3 more reasoning resultsHide 3 reasoning results

Coding

Not enough dataFull coding ranking

0 of 10 ranked benchmarks measured

Show 1 more coding resultHide 1 coding result

More results

Benchmarks outside the capability baskets. They are not ranked against a leader.

Show 4 more resultsHide 4 results

Gemini 2.5 Flash (Sep) (Non-Reasoning): common questions

Who makes Gemini 2.5 Flash (Sep) (Non-Reasoning)?

Gemini 2.5 Flash (Sep) (Non-Reasoning) is made by Google.

When was Gemini 2.5 Flash (Sep) (Non-Reasoning) released?

Gemini 2.5 Flash (Sep) (Non-Reasoning) was released on Sep 25, 2025, according to Artificial Analysis.

What is Gemini 2.5 Flash (Sep) (Non-Reasoning) good at?

Gemini 2.5 Flash (Sep) (Non-Reasoning) is capable in long context; and behind the leaders in multimodal tasks, factuality, agentic tasks, instruction following, and reasoning. Too few results yet to rate coding, safety, math, or multilingual tasks.

How many benchmarks has Gemini 2.5 Flash (Sep) (Non-Reasoning) been tested on?

We track 34 results for Gemini 2.5 Flash (Sep) (Non-Reasoning) on 17 benchmarks from 3 sources, 2 of them independently verified. The latest was recorded on Oct 8, 2026.

About this record

Where Gemini 2.5 Flash (Sep) (Non-Reasoning)'s numbers come from, and every name it appears under.

Tracked since
Sep 11, 2026
Newest source mention
Sep 12, 2026

Where the results come from

Verification: 34 scores · 2 independently verified · 32 aggregator-attributed. How these tiers are assigned

From 3 sources on 3 sites. Artificial Analysis supplies 32 of them; the 2 independently verified results come from 2 sites. Bars are coloured by trust tier.

  • artificialanalysis.ai32
  • datasets-server.huggingface.co1
  • lmarena.ai1

Also known as

Gemini 2.5 Flash Preview 09-2025gemini 2.5 flash (sep)