VECTOR WIREAI INTELLIGENCE
PKT
Refresh Models Deals Regulatory Sources

Questions this page answers

Which is better, Claude Opus 4.5 or Qwen3.5 122B A10B?
Across 51 shared benchmarks, Claude Opus 4.5 scores higher on 31 and Qwen3.5 122B A10B on 20. The widest gap is critpt, where Claude Opus 4.5 scores 4.6 against 0.9. Qwen3.5 122B A10B is the cheaper of the two on tracked API pricing ($0.40 against $5.00 per million input tokens).

Claude Opus 4.5 vs Qwen3.5 122B A10B

Across 51 shared benchmarks, Claude Opus 4.5 scores higher on 31 and Qwen3.5 122B A10B on 20. The widest gap is critpt, where Claude Opus 4.5 scores 4.6 against 0.9. Qwen3.5 122B A10B is the cheaper of the two on tracked API pricing ($0.40 against $5.00 per million input tokens).

AnthropicvsAlibaba51 shared benchmarks3120 head-to-head
BenchmarkClaude Opus 4.5Qwen3.5 122B A10B
AA Agentic Index59.621.3
AA Intelligence41.932.8
AA-LCR7670.3
AA-Omniscience14-41.5
Artificial Analysis Coding Index47.845.7
browsecomp3763.8
browsecomp_zh62.469.9
C-Eval92.291.9
CC-OCR76.981.8
coding_arena_elo14951358
critpt4.60.9
DynaMath79.785.9
ERQA46.862
gdpval47.324.3
GPQA Diamond8786.6
HLE30.847.5
IFBench5876.1
LiveCodeBench v684.878.9
longbench_v264.460.2
mathvision77.186.2
MathVista80.287.4
MLVU81.787.3
mmlu_redux95.694
MMLU-Pro9086.7
mmmlu90.886.7
MMMU80.783.9
MMMU-Pro7476.9
MMStar73.282.9
MMVU77.374.7
multichallenge5961.5
MV-Bench67.276.6
OCRBench86.592.1
OmniDocBench 1.587.789.8
OmniScience Accuracy46.624.4
OmniScience Non-Hallucination3912.9
OSWorld-Verified66.358
RealWorldQA7785.1
scicode5042
Seal-047.744.1
SimpleVQA69.70.6
supergpqa70.667.1
SWE-bench Verified81.572
Terminal-Bench 2.059.349.4
Terminal-Bench Hard4731.1
vectara_answer_rate98.799.8
vectara_avg_summary_length114.586.4
vectara_factual_consistency89.188.8
vectara_hallucination_rate10.911.2
VideoMMMU84.482
ZeroBench30.1
τ²-Bench Telecom (AA run)89.593.6

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing. Quoted rates are the price-setter row we currently track for each model — its direct or vendor-official listing where one exists (anthropic-official, direct), otherwise the lowest tracked offer.