Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

Qwen3.8 27B

Score basis

Qwen3.8 27B is strong in long context; capable in instruction following, agentic tasks, coding, multimodal tasks, and reasoning; and behind the leaders in factuality and math. Too few results yet to rate safety or multilingual tasks.

Capability profile

Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.

Strong1 capability
  1. Long Context−8.4%1 of 3
Capable5 capabilities
  1. Instruction Following−14.1%2 of 3
  2. Agentic−16.4%4 of 7
  3. Coding−18.5%8 of 10
  4. Multimodal−19.2%2 of 6
  5. Reasoning−24.2%6 of 6
Limited2 capabilities
  1. Factuality−31.8%2 of 4
  2. Math−40.5%1 of 5
Not rated2 capabilities

Too few results yet to rate Safety or Multilingual.

Price

$0.50input$3.00outputper million tokens

From Alibaba's own price page · 7 providers tracked · All prices

Costs more than 61% of 331 priced models · 3:1 input-to-output blend, log scale

Evidence

108results on52benchmarks

  • 20 independently verified
  • 62 aggregator
  • 26 vendor-reported

From 12 sources · latest Oct 8, 2026 · How verification works

API features

Tool callingStructured outputsReasoning

As listed by OpenRouter

Qwen3.8 27B benchmark results

108 results on 52 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.

Long Context

StrongFull long context ranking

8.4% behind the leader1 of 3 ranked benchmarks measured

Show 3 more long context resultsHide 3 long context results

Instruction Following

CapableFull instruction following ranking

14.1% behind the leader2 of 3 ranked benchmarks measured

Agentic

CapableFull agentic ranking

16.4% behind the leader4 of 7 ranked benchmarks measured

Show 9 more agentic resultsHide 9 agentic results

Coding

CapableFull coding ranking

18.5% behind the leader8 of 10 ranked benchmarks measured

Show 7 more coding resultsHide 7 coding results

Multimodal

CapableFull multimodal ranking

19.2% behind the leader2 of 6 ranked benchmarks measured

Show 5 more multimodal resultsHide 5 multimodal results

Reasoning

CapableFull reasoning ranking

24.2% behind the leader6 of 6 ranked benchmarks measured

Show 12 more reasoning resultsHide 12 reasoning results

Factuality

LimitedFull factuality ranking

31.8% behind the leader2 of 4 ranked benchmarks measured

Show 6 more factuality resultsHide 6 factuality results

40.5% behind the leader1 of 5 ranked benchmarks measured

More results

Benchmarks outside the capability baskets. They are not ranked against a leader.

Show 34 more resultsHide 34 results

Qwen3.8 27B: common questions

Who makes Qwen3.8 27B?

Qwen3.8 27B is made by Alibaba.

When was Qwen3.8 27B released?

Qwen3.8 27B was released on Aug 14, 2026, according to Artificial Analysis.

What is Qwen3.8 27B good at?

Qwen3.8 27B is strong in long context; capable in instruction following, agentic tasks, coding, multimodal tasks, and reasoning; and behind the leaders in factuality and math. Too few results yet to rate safety or multilingual tasks.

How much does Qwen3.8 27B cost?

Qwen3.8 27B costs $0.50 per million input tokens and $3.00 per million output tokens, according to Alibaba's own price page. We track its price at 7 providers. At a mix of three input tokens to one output token, it costs more than 61% of the 331 priced models we track.

How many benchmarks has Qwen3.8 27B been tested on?

We track 108 results for Qwen3.8 27B on 52 benchmarks from 12 sources, 20 of them independently verified. The latest was recorded on Oct 8, 2026.

Which API features does Qwen3.8 27B support?

OpenRouter lists tool calling, structured outputs, and reasoning for Qwen3.8 27B.

About this record

Where Qwen3.8 27B's numbers come from, and every name it appears under.

Tracked since
Aug 3, 2026
Newest source mention
Oct 2, 2026

Where the results come from

Verification: 108 scores · 20 independently verified · 62 aggregator-attributed · 26 vendor-reported. How these tiers are assigned

From 12 sources on 8 sites. Artificial Analysis supplies 62 of them; the 20 independently verified results come from 5 sites. Bars are coloured by trust tier.

  • artificialanalysis.ai62
  • api.llm-stats.com20
  • arcprize.org8
  • livebench.ai7
  • huggingface.co6
  • 99franklin.github.io2
  • datasets-server.huggingface.co2
  • simple-bench.com1

Also known as

How our sources name Qwen3.8 27B at each reasoning setting.

SettingShort formAPI id
lowqwen3.8 27b (low)qwen3-8-27b-low
mediumqwen3.8 27b (medium)qwen3-8-27b-medium
xhighqwen3.8 27b (xhigh)—
Also listed asqwen 3.8 - 27bqwen/qwen3.8-27bqwen3.8-27b q8_0qwen3-8-27bqwen3.8:27bqwen3.8 27b (non-reasoning)qwen3-8-27b-non-reasoningqwen 3.8 27b denseqwen3.8-27b (none)