VECTOR WIREAI INTELLIGENCE
NVDA$1,847+3.2%MSFT$512+1.1%GOOGL$199-0.4%META$728+2.7%AMD$184-1.2%TSM$212+0.6%PLTR$98+4.1%AI IDX4,821+1.9%
PKT
SEEDRefresh Models Deals Regulatory Sources

GLM-5.2 Full Open Source vs Kimi K2.5

Z.aivsMoonshot46 shared benchmarks406 head-to-head
BenchmarkGLM-5.2 Full Open SourceKimi K2.5
AA Agentic Index45.752.8
AA Intelligence52.636
AA-LCR76.773
AA-Omniscience4.4-7.3
aime_202699.295.8
apexAgents33.711.5
arc_agi_17765.3
ARC-AGI-222.811.8
arena_elo14711450
arena_text_factuality14601445
Artificial Analysis Coding Index68.846.8
coding_arena_elo15821436
critpt20.93.1
FORTRESS (Adversarial)71.354.1
FORTRESS (Benign)9098.3
gdpval50.338.3
GDPval-AA v215141009
Global-MMLU-Lite89.284
GPQA Diamond91.287.9
HLE54.750.2
HLE (with tools)54.751.8
HMMT 202594.495.4
hmmt_feb_202692.587.1
hmmt_nov_202594.491.1
IFBench73.370.2
imo_answer_bench9181.8
MCP Atlas82.664
nl2repo48.932
OmniScience Accuracy24.335.2
OmniScience Non-Hallucination73.750
ResearchRubrics71.159.5
scicode50.549
simplebench58.846.8
simpleqa_verified38.136.9
strongreject98.599.5
SWE-bench Pro62.153.8
SWEBench Pro (Public)62.150.7
Tau 3 Banking26.814.2
TauBench V3 - Banking34.614.2
Terminal Bench 2.1 (Best Harness)82.751.3
Terminal-Bench 2.182.745.7
Terminal-Bench Hard50.834.8
tool_decathlon48.227.8
toolathlon48.227.8
τ²-Bench Telecom (AA run)99.195.9
τ³-Bench26.866

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.