Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

Qwen3 4B 2507 Instruct

Score basis

Qwen3 4B 2507 Instruct is behind the leaders in agentic tasks. Too few results yet to rate reasoning, coding, safety, long context, math, multimodal tasks, multilingual tasks, instruction following, or factuality.

Capability profile

Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.

Limited1 capability
  1. Agentic−47.0%1 of 7
Not rated9 capabilities

Too few results yet to rate Reasoning, Coding, Safety, Long Context, Math, Multimodal, Multilingual, Instruction Following or Factuality.

Price

$0.01input$0.03outputper million tokens

From nscale · All prices

Cheaper than 99% of 329 priced models · 3:1 input-to-output blend, log scale

Evidence

29results on27benchmarks

  • 15 aggregator
  • 14 cross-referenced

From 3 sources · latest Oct 7, 2026 · How verification works

Qwen3 4B 2507 Instruct benchmark results

29 results on 27 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.

Agentic

LimitedFull agentic ranking

47.0% behind the leader1 of 7 ranked benchmarks measured

Reasoning

Not enough dataFull reasoning ranking

0 of 6 ranked benchmarks measured

Show 1 more reasoning resultHide 1 reasoning result

Coding

Not enough dataFull coding ranking

0 of 10 ranked benchmarks measured

Long Context

Not enough dataFull long context ranking

0 of 3 ranked benchmarks measured

Instruction Following

Not enough dataFull instruction following ranking

0 of 3 ranked benchmarks measured

Factuality

Not enough dataFull factuality ranking

0 of 4 ranked benchmarks measured

More results

Benchmarks outside the capability baskets. They are not ranked against a leader.

Show 10 more resultsHide 10 results

Qwen3 4B 2507 Instruct: common questions

Who makes Qwen3 4B 2507 Instruct?

Qwen3 4B 2507 Instruct is made by Alibaba.

When was Qwen3 4B 2507 Instruct released?

Qwen3 4B 2507 Instruct was released on Aug 6, 2025, according to Artificial Analysis.

What is Qwen3 4B 2507 Instruct good at?

Qwen3 4B 2507 Instruct is behind the leaders in agentic tasks. Too few results yet to rate reasoning, coding, safety, long context, math, multimodal tasks, multilingual tasks, instruction following, or factuality.

How much does Qwen3 4B 2507 Instruct cost?

Qwen3 4B 2507 Instruct costs $0.01 per million input tokens and $0.03 per million output tokens, according to nscale. At a mix of three input tokens to one output token, it is cheaper than 99% of the 329 priced models we track.

How many benchmarks has Qwen3 4B 2507 Instruct been tested on?

We track 29 results for Qwen3 4B 2507 Instruct on 27 benchmarks from 3 sources. The latest was recorded on Oct 7, 2026.

About this record

Where Qwen3 4B 2507 Instruct's numbers come from, and every name it appears under.

Tracked since
May 2, 2026
Newest source mention
Sep 10, 2026

Where the results come from

Verification: 29 scores · 0 independently verified · 15 aggregator-attributed · 14 vendor cross-reference. How these tiers are assigned

From 3 sources on 2 sites. Artificial Analysis supplies 15 of them. Bars are coloured by trust tier.

  • artificialanalysis.ai15
  • huggingface.co14

Also known as

qwen3 4b 2507 (non-reasoning)Qwen3-4B-Instruct-2507