VECTOR WIREAI INTELLIGENCE
PKT
Refresh Models Deals Regulatory Sources

Questions this page answers

Which is better, Claude Sonnet 5 or Qwen3.5 122B A10B?
Across 23 shared benchmarks, Claude Sonnet 5 scores higher on 22 and Qwen3.5 122B A10B on 1. The widest gap is OmniScience Non-Hallucination, where Claude Sonnet 5 scores 60.6 against 12.9. Qwen3.5 122B A10B is the cheaper of the two on tracked API pricing ($0.40 against $2.00 per million input tokens).

Claude Sonnet 5 vs Qwen3.5 122B A10B

Across 23 shared benchmarks, Claude Sonnet 5 scores higher on 22 and Qwen3.5 122B A10B on 1. The widest gap is OmniScience Non-Hallucination, where Claude Sonnet 5 scores 60.6 against 12.9. Qwen3.5 122B A10B is the cheaper of the two on tracked API pricing ($0.40 against $2.00 per million input tokens).

AnthropicvsAlibaba23 shared benchmarks221 head-to-head
BenchmarkClaude Sonnet 5Qwen3.5 122B A10B
AA Agentic Index49.721.3
AA Intelligence55.332.8
AA-LCR7770.3
AA-Omniscience16.4-41.5
arena_vision12811246
Artificial Analysis Coding Index71.545.7
browsecomp84.763.8
coding_arena_elo15391358
critpt16.90.9
gdpval54.624.3
GPQA Diamond91.186.6
HLE57.447.5
LVBench68.574.4
MMMU-Pro77.376.9
OmniScience Accuracy4024.4
OmniScience Non-Hallucination60.612.9
OSWorld-Verified81.258
scicode53.642
SWE-bench Verified85.272
TauBench V3 - Banking37.315.3
Terminal-Bench 2.080.449.4
Terminal-Bench 2.180.547.6
τ³-Bench28.213.6

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing. Quoted rates are the price-setter row we currently track for each model — its direct or vendor-official listing where one exists (anthropic-official, direct), otherwise the lowest tracked offer.