VECTOR WIREAI INTELLIGENCE
NVDA$1,847+3.2%MSFT$512+1.1%GOOGL$199-0.4%META$728+2.7%AMD$184-1.2%TSM$212+0.6%PLTR$98+4.1%AI IDX4,821+1.9%
PKT
SEEDRefresh Models Deals Regulatory Sources

GLM-5.2 Full Open Source vs Kimi K2.6

Z.aivsMoonshot48 shared benchmarks3512 head-to-head
BenchmarkGLM-5.2 Full Open SourceKimi K2.6
AA Agentic Index45.758.7
AA Intelligence52.645.1
AA-LCR76.776.7
AA-Omniscience4.46.4
aime_202699.296.4
apex_agents35.627.9
apexAgents33.728.5
arena_elo14711461
arena_text_factuality14601455
Artificial Analysis Coding Index68.861.8
coding_arena_elo15821509
critpt20.98
FORTRESS (Adversarial)71.365.6
FORTRESS (Benign)9097.2
gdpval50.341.4
GDPval-AA v215141190
Global-MMLU-Lite89.288.4
GPQA Diamond91.291.1
HiL-Bench43.718.7
HLE54.737.5
HLE (with tools)54.754
hmmt_feb_202692.594.7
IFBench73.376
imo_answer_bench9186
itbenchSre42.731.2
livebench_agentic_coding51.846.9
livebench_coding79.778.6
livebench_data_analysis73.765.1
livebench_instruction_following62.364.4
livebench_language76.275.1
livebench_math89.884.3
livebench_reasoning78.679.4
MCP Atlas82.668.1
OmniScience Accuracy24.332.8
OmniScience Non-Hallucination73.760.7
scicode50.553.5
simpleqa_verified38.138.7
strongreject98.599.8
SWE-bench Pro62.158.6
SWEBench Pro (Public)62.158.6
Tau 3 Banking26.820.6
TauBench V3 - Banking34.623.3
Terminal Bench 2.1 (Best Harness)82.771.3
Terminal-Bench 2.182.765.9
Terminal-Bench Hard50.843.9
toolathlon48.250
τ²-Bench Telecom (AA run)99.195.9
τ³-Bench26.820.6

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.