Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

Qwen3 VL 4B (Reasoning)

Score basis

Qwen3 VL 4B (Reasoning) is behind the leaders in agentic tasks and multimodal tasks. Too few results yet to rate reasoning, coding, safety, long context, math, multilingual tasks, instruction following, or factuality.

Capability profile

Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.

Limited2 capabilities
  1. Agentic−45.1%2 of 7
  2. Multimodal−54.8%5 of 6
Not rated8 capabilities

Too few results yet to rate Reasoning, Coding, Safety, Long Context, Math, Multilingual, Instruction Following or Factuality.

Price

No current price is tracked for this model. See the rate card

Evidence

118results on102benchmarks

  • 16 aggregator
  • 37 vendor-reported
  • 65 cross-referenced

From 6 sources · latest Oct 8, 2026 · How verification works

Qwen3 VL 4B (Reasoning) benchmark results

118 results on 102 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.

Agentic

LimitedFull agentic ranking

45.1% behind the leader2 of 7 ranked benchmarks measured

Multimodal

LimitedFull multimodal ranking

54.8% behind the leader5 of 6 ranked benchmarks measured

Show 7 more multimodal resultsHide 7 multimodal results

Reasoning

Not enough dataFull reasoning ranking

0 of 6 ranked benchmarks measured

Show 2 more reasoning resultsHide 2 reasoning results

Coding

Not enough dataFull coding ranking

0 of 10 ranked benchmarks measured

Long Context

Not enough dataFull long context ranking

0 of 3 ranked benchmarks measured

Instruction Following

Not enough dataFull instruction following ranking

0 of 3 ranked benchmarks measured

Factuality

Not enough dataFull factuality ranking

0 of 4 ranked benchmarks measured

More results

Benchmarks outside the capability baskets. They are not ranked against a leader.

Show 86 more resultsHide 86 results

Qwen3 VL 4B (Reasoning): common questions

Who makes Qwen3 VL 4B (Reasoning)?

Qwen3 VL 4B (Reasoning) is made by Alibaba.

When was Qwen3 VL 4B (Reasoning) released?

Qwen3 VL 4B (Reasoning) was released on Oct 14, 2025, according to Artificial Analysis.

What is Qwen3 VL 4B (Reasoning) good at?

Qwen3 VL 4B (Reasoning) is behind the leaders in agentic tasks and multimodal tasks. Too few results yet to rate reasoning, coding, safety, long context, math, multilingual tasks, instruction following, or factuality.

How many benchmarks has Qwen3 VL 4B (Reasoning) been tested on?

We track 118 results for Qwen3 VL 4B (Reasoning) on 102 benchmarks from 6 sources. The latest was recorded on Oct 8, 2026.

About this record

Where Qwen3 VL 4B (Reasoning)'s numbers come from, and every name it appears under.

Tracked since
May 2, 2026
Newest source mention
Sep 10, 2026

Where the results come from

Verification: 118 scores · 0 independently verified · 16 aggregator-attributed · 65 vendor cross-reference · 37 vendor-reported. How these tiers are assigned

From 6 sources on 3 sites. Hugging Face supplies 65 of them. Bars are coloured by trust tier.

  • huggingface.co65
  • api.llm-stats.com37
  • artificialanalysis.ai16

Also known as

qwen3-vl-4b-thinkingqwen3-vl 4b