VECTOR WIREAI INTELLIGENCE
NVDA$1,847+3.2%MSFT$512+1.1%GOOGL$199-0.4%META$728+2.7%AMD$184-1.2%TSM$212+0.6%PLTR$98+4.1%AI IDX4,821+1.9%
PKT
SEEDRefresh Models Deals Regulatory Sources

GLM-5.2 Full Open Source vs Qwen3.6 27B

Z.aivsAlibaba35 shared benchmarks341 head-to-head
BenchmarkGLM-5.2 Full Open SourceQwen3.6 27B
AA Agentic Index45.727.5
AA Intelligence52.637.7
AA-LCR76.773.3
AA-Omniscience4.4-20
Agents' Last Exam23.827.3
aime_202699.294.1
Artificial Analysis Coding Index68.853.7
critpt20.91.1
DeepSWE 1.14413.3
gdpval50.332
GPQA Diamond91.287.8
HLE54.724
HMMT 202594.490.7
hmmt_feb_202692.584.3
hmmt_nov_202594.490.7
IFBench73.369.1
imo_answer_bench9180.8
JobBench43.421.8
livebench_agentic_coding51.839.3
livebench_coding79.771.8
livebench_data_analysis73.770.4
livebench_instruction_following62.353.2
livebench_language76.263.3
livebench_math89.879.9
livebench_reasoning78.670.3
nl2repo48.936.2
OmniScience Accuracy24.319.6
OmniScience Non-Hallucination73.750.7
scicode50.539.8
SWE-bench Pro62.153.5
TauBench V3 - Banking34.616.7
Terminal-Bench 2.182.760.7
Terminal-Bench Hard50.834.8
τ²-Bench Telecom (AA run)99.194.2
τ³-Bench26.815.3

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.