VECTOR WIREAI INTELLIGENCE
PKT
Refresh Models Deals Regulatory Sources

Questions this page answers

Which is better, Qwen2.5 Instruct 72B or Qwen3.5 122B A10B?
Across 21 shared benchmarks, Qwen2.5 Instruct 72B scores higher on 2 and Qwen3.5 122B A10B on 19. The widest gap is Terminal-Bench Hard, where Qwen3.5 122B A10B scores 31.1 against 4.5. Tracked API pricing per million tokens: Qwen2.5 Instruct 72B $0.47 in / $0.49 out, Qwen3.5 122B A10B $0.40 in / $3.20 out.

Qwen2.5 Instruct 72B vs Qwen3.5 122B A10B

Across 21 shared benchmarks, Qwen2.5 Instruct 72B scores higher on 2 and Qwen3.5 122B A10B on 19. The widest gap is Terminal-Bench Hard, where Qwen3.5 122B A10B scores 31.1 against 4.5. Tracked API pricing per million tokens: Qwen2.5 Instruct 72B $0.47 in / $0.49 out, Qwen3.5 122B A10B $0.40 in / $3.20 out.

AlibabavsAlibaba21 shared benchmarks219 head-to-head
BenchmarkQwen2.5 Instruct 72BQwen3.5 122B A10B
AA Intelligence1032.8
AA-LCR20.370.3
AA-Omniscience-52.2-41.5
Artificial Analysis Coding Index11.945.7
C-Eval89.291.9
critpt00.9
GPQA Diamond49.186.6
GSM8K95.894.5
HLE4.247.5
IFBench36.976.1
ifeval87.293.4
longbench_v239.460.2
mmlu_redux86.894
MMLU-Pro71.686.7
mmmlu74.886.7
OmniScience Accuracy17.624.4
OmniScience Non-Hallucination15.312.9
scicode26.742
SWE-bench Verified23.872
Terminal-Bench Hard4.531.1
τ²-Bench Telecom (AA run)34.593.6

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing. Quoted rates are the price-setter row we currently track for each model — its direct or vendor-official listing where one exists (direct, direct), otherwise the lowest tracked offer.