VECTOR WIREAI INTELLIGENCE
NVDA$1,847+3.2%MSFT$512+1.1%GOOGL$199-0.4%META$728+2.7%AMD$184-1.2%TSM$212+0.6%PLTR$98+4.1%AI IDX4,821+1.9%
PKT
SEEDRefresh Models Deals Regulatory Sources

Qwen2.5 Instruct 72B vs Qwen3.5 35B A3B

AlibabavsAlibaba28 shared benchmarks721 head-to-head
BenchmarkQwen2.5 Instruct 72BQwen3.5 35B A3B
AA Intelligence1029.9
AA-LCR20.368.3
AA-Omniscience-52.2-48.1
arc_challenge94.595.4
Artificial Analysis Coding Index11.937
C-Eval89.290.2
critpt00.9
GPQA Diamond49.184.5
GSM8K95.890.1
hellaswag84.885.6
HLE4.247.4
humaneval86.666.5
IFBench36.972.5
ifeval87.291.9
longbench_v239.459
MBPP72.670.8
mmlu86.181.1
mmlu_redux86.893.3
MMLU-Pro71.685.3
mmmlu74.885.2
OmniScience Accuracy17.620.1
OmniScience Non-Hallucination15.314.6
OpenBookQA96.244.2
scicode26.737.7
SWE-bench Verified23.870
Terminal-Bench Hard4.526.5
winogrande82.379.2
τ²-Bench Telecom (AA run)34.589.2

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.