VECTOR WIREAI INTELLIGENCE
NVDA$1,847+3.2%MSFT$512+1.1%GOOGL$199-0.4%META$728+2.7%AMD$184-1.2%TSM$212+0.6%PLTR$98+4.1%AI IDX4,821+1.9%
PKT
SEEDRefresh Models Deals Regulatory Sources

Gemma 4 26B A4B vs Qwen3.5 397B A17B

GooglevsAlibaba59 shared benchmarks356 head-to-head
BenchmarkGemma 4 26B A4BQwen3.5 397B A17B
AA Agentic Index32.153.3
AA Intelligence26.134.3
AA-LCR61.772.7
AA-Omniscience-50.8-30.8
aime_202688.394.2
arena_elo14381442
arena_text_factuality14451445
arena_vision12381249
Artificial Analysis Coding Index39.348.2
browsecomp26.369
C-Eval82.593
CC-OCR74.582
charxiv_rq6980.8
Claw Eval (pass@3)2848.1
Claw-Eval Avg58.870.7
coding_arena_elo13611399
critpt01.7
DeepPlanning16.234.3
gdpval25.735.8
GDPval-AA v2807962
GPQA Diamond82.389.3
GSM8K7766.7
HLE19.329
HLE (with tools)17.248.3
hmmt_feb_202591.794.8
hmmt_feb_20267987.9
hmmt_nov_202587.592.7
IFBench72.478.8
ifeval96.492.6
imo_answer_bench74.380.9
itbenchSre23.634.1
LiveCodeBench v677.183.6
mcpmark14.246.1
mmlu_redux92.794.9
MMLU-Pro85.287.8
mmmlu86.388.5
MMMU78.485
MMMU-Pro73.879
nl2repo11.632.2
OmniScience Accuracy19.130.8
OmniScience Non-Hallucination13.617.3
QwenClawBench38.751.8
QwenWebBench11781186
RealWorldQA72.283.9
scicode40.342
SimpleVQA52.267.1
SkillsBench Avg512.330
supergpqa61.470.4
SWE-bench Multilingual43.469.3
SWE-bench Pro13.850.9
SWE-bench Verified57.476.4
TauBench V3 - Banking1213.4
Terminal-Bench 2.034.252.5
Terminal-Bench 2.13951.3
Terminal-Bench Hard2540.9
VITA-Bench36.949.7
WideSearch38.374
τ²-Bench Telecom (AA run)43.695.6
τ³-Bench5913.4

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.