Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

Qwen3 235B A22B Instruct 2507

Score basis

Qwen3 235B A22B Instruct 2507 is behind the leaders in factuality, instruction following, and agentic tasks. Too few results yet to rate reasoning, coding, safety, long context, math, multimodal tasks, or multilingual tasks.

Capability profile

Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.

Limited3 capabilities
  1. Factuality−36.7%2 of 4
  2. Instruction Following−40.5%1 of 3
  3. Agentic−40.7%1 of 7
Not rated7 capabilities

Too few results yet to rate Reasoning, Coding, Safety, Long Context, Math, Multimodal or Multilingual.

Price

$0.23input$0.92outputper million tokens

From Alibaba's own price page · 6 providers tracked · All prices

Cheaper than 65% of 330 priced models · 3:1 input-to-output blend, log scale

Evidence

41results on40benchmarks

  • 8 independently verified
  • 15 aggregator
  • 18 vendor-reported

From 10 sources · latest Oct 8, 2026 · How verification works

Qwen3 235B A22B Instruct 2507 benchmark results

41 results on 40 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.

Factuality

LimitedFull factuality ranking

36.7% behind the leader2 of 4 ranked benchmarks measured

Instruction Following

LimitedFull instruction following ranking

40.5% behind the leader1 of 3 ranked benchmarks measured

Agentic

LimitedFull agentic ranking

40.7% behind the leader1 of 7 ranked benchmarks measured

Reasoning

Not enough dataFull reasoning ranking

0 of 6 ranked benchmarks measured

Show 2 more reasoning resultsHide 2 reasoning results

Coding

Not enough dataFull coding ranking

0 of 10 ranked benchmarks measured

Long Context

Not enough dataFull long context ranking

0 of 3 ranked benchmarks measured

More results

Benchmarks outside the capability baskets. They are not ranked against a leader.

Show 22 more resultsHide 22 results

Qwen3 235B A22B Instruct 2507: common questions

Who makes Qwen3 235B A22B Instruct 2507?

Qwen3 235B A22B Instruct 2507 is made by Alibaba.

When was Qwen3 235B A22B Instruct 2507 released?

Qwen3 235B A22B Instruct 2507 was released on Jul 21, 2025, according to Artificial Analysis.

What is Qwen3 235B A22B Instruct 2507 good at?

Qwen3 235B A22B Instruct 2507 is behind the leaders in factuality, instruction following, and agentic tasks. Too few results yet to rate reasoning, coding, safety, long context, math, multimodal tasks, or multilingual tasks.

How much does Qwen3 235B A22B Instruct 2507 cost?

Qwen3 235B A22B Instruct 2507 costs $0.23 per million input tokens and $0.92 per million output tokens, according to Alibaba's own price page. We track its price at 6 providers. At a mix of three input tokens to one output token, it is cheaper than 65% of the 330 priced models we track.

How many benchmarks has Qwen3 235B A22B Instruct 2507 been tested on?

We track 41 results for Qwen3 235B A22B Instruct 2507 on 40 benchmarks from 10 sources, 8 of them independently verified. The latest was recorded on Oct 8, 2026.

About this record

Where Qwen3 235B A22B Instruct 2507's numbers come from, and every name it appears under.

Tracked since
Apr 25, 2026
Newest source mention
May 19, 2026

Where the results come from

Verification: 41 scores · 8 independently verified · 15 aggregator-attributed · 18 vendor-reported. How these tiers are assigned

From 10 sources on 4 sites. api.llm-stats.com supplies 18 of them; the 8 independently verified results come from 2 sites. Bars are coloured by trust tier.

  • api.llm-stats.com18
  • artificialanalysis.ai15
  • storage.googleapis.com6
  • arcprize.org2

Also known as

qwen3-235b-a22b-instruct-2507 (Reasoning)qwen3-235b-a22b instruct (25/07)qwen3-235b-a22b-instruct-reasoningqwen3-235b-a22b-instructqwen3 235b a22b 2507 instructqwen3 235b a22b instruct 2507 fp8qwen3-235b-a22b-instruct-2507-reasoningqwen3-235b-a22b-instruct-2507 (non-reasoning)