Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

Qwen3 VL Thinking (8B)

Score basis

Qwen3 VL Thinking (8B) is behind the leaders in multimodal tasks and agentic tasks. Too few results yet to rate reasoning, coding, safety, long context, math, multilingual tasks, instruction following, or factuality.

Capability profile

Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.

Limited2 capabilities
  1. Multimodal−45.8%6 of 6
  2. Agentic−46.0%2 of 7
Not rated8 capabilities

Too few results yet to rate Reasoning, Coding, Safety, Long Context, Math, Multilingual, Instruction Following or Factuality.

Price

$0.18input$2.10outputper million tokens

From Alibaba's own price page · 3 providers tracked · All prices

Cheaper than 52% of 331 priced models · 3:1 input-to-output blend, log scale

Evidence

74results on63benchmarks

  • 16 aggregator
  • 39 vendor-reported
  • 19 cross-referenced

From 4 sources · latest Oct 8, 2026 · How verification works

API features

Tool callingStructured outputsReasoning

As listed by OpenRouter

Qwen3 VL Thinking (8B) benchmark results

74 results on 63 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.

Multimodal

LimitedFull multimodal ranking

45.8% behind the leader6 of 6 ranked benchmarks measured

Show 7 more multimodal resultsHide 7 multimodal results

Agentic

LimitedFull agentic ranking

46.0% behind the leader2 of 7 ranked benchmarks measured

Reasoning

Not enough dataFull reasoning ranking

0 of 6 ranked benchmarks measured

Show 2 more reasoning resultsHide 2 reasoning results

Coding

Not enough dataFull coding ranking

0 of 10 ranked benchmarks measured

Long Context

Not enough dataFull long context ranking

0 of 3 ranked benchmarks measured

Instruction Following

Not enough dataFull instruction following ranking

0 of 3 ranked benchmarks measured

Factuality

Not enough dataFull factuality ranking

0 of 4 ranked benchmarks measured

More results

Benchmarks outside the capability baskets. They are not ranked against a leader.

Show 41 more resultsHide 41 results

Qwen3 VL Thinking (8B): common questions

Who makes Qwen3 VL Thinking (8B)?

Qwen3 VL Thinking (8B) is made by Alibaba.

When was Qwen3 VL Thinking (8B) released?

Qwen3 VL Thinking (8B) was released on Oct 14, 2025, according to Artificial Analysis.

What is Qwen3 VL Thinking (8B) good at?

Qwen3 VL Thinking (8B) is behind the leaders in multimodal tasks and agentic tasks. Too few results yet to rate reasoning, coding, safety, long context, math, multilingual tasks, instruction following, or factuality.

How much does Qwen3 VL Thinking (8B) cost?

Qwen3 VL Thinking (8B) costs $0.18 per million input tokens and $2.10 per million output tokens, according to Alibaba's own price page. We track its price at 3 providers. At a mix of three input tokens to one output token, it is cheaper than 52% of the 331 priced models we track.

How many benchmarks has Qwen3 VL Thinking (8B) been tested on?

We track 74 results for Qwen3 VL Thinking (8B) on 63 benchmarks from 4 sources. The latest was recorded on Oct 8, 2026.

Which API features does Qwen3 VL Thinking (8B) support?

OpenRouter lists tool calling, structured outputs, and reasoning for Qwen3 VL Thinking (8B).

About this record

Where Qwen3 VL Thinking (8B)'s numbers come from, and every name it appears under.

Tracked since
May 16, 2026
Newest source mention
Sep 29, 2026

Where the results come from

Verification: 74 scores · 0 independently verified · 16 aggregator-attributed · 19 vendor cross-reference · 39 vendor-reported. How these tiers are assigned

From 4 sources on 3 sites. api.llm-stats.com supplies 39 of them. Bars are coloured by trust tier.

  • api.llm-stats.com39
  • huggingface.co19
  • artificialanalysis.ai16

Also known as

qwen3-vl-8b-reasoningqwen3-vl-8bqwen3-vl-8b-thinkingqwen3 vl 8b (reasoning)qwen3-vl:8b