VECTOR WIREAI INTELLIGENCE
NVDA$1,847+3.2%MSFT$512+1.1%GOOGL$199-0.4%META$728+2.7%AMD$184-1.2%TSM$212+0.6%PLTR$98+4.1%AI IDX4,821+1.9%
PKT
SEEDRefresh Models Deals Regulatory Sources

DeepSeek-V4-Pro vs GLM-5.2 Full Open Source

DeepSeekvsZ.ai58 shared benchmarks1740 head-to-head
BenchmarkDeepSeek-V4-ProGLM-5.2 Full Open Source
AA Agentic Index63.345.7
AA Intelligence5352.6
AA-LCR7076.7
AA-Omniscience-10.64.4
Agents' Last Exam16.523.8
aime_202696.799.2
apexAgents24.333.7
arena_elo14571471
arena_text_factuality14501460
Artificial Analysis Coding Index59.468.8
AutomationBench (Public)12.812.9
AutomationBench Public12.812.9
coding_arena_elo15821582
critpt1320.9
DeepSWE12.846.2
DSBench-FullStack41.861.8
DSBench-Hard31.154.5
FORTRESS (Adversarial)3671.3
FORTRESS (Benign)98.590
gdpval4950.3
GDPval-AA v213071514
Global-MMLU-Lite89.389.2
GPQA Diamond90.591.2
HLE48.254.7
HLE (with tools)48.254.7
HLE (wo / w tools)48.254.7
hmmt_feb_202695.292.5
hmmt_nov_202594.494.4
IFBench76.573.3
imo_answer_bench89.891
itbenchSre38.342.7
livebench_agentic_coding42.651.8
livebench_coding7079.7
livebench_data_analysis74.573.7
livebench_instruction_following62.462.3
livebench_language78.176.2
livebench_math90.789.8
livebench_reasoning82.778.6
MCP Atlas74.282.6
nl2repo38.548.9
OmniScience Accuracy4324.3
OmniScience Non-Hallucination12.273.7
Program Bench47.863.7
scicode5050.5
simplebench50.958.8
simpleqa_verified57.938.1
strongreject98.698.5
SWE-bench Pro55.462.1
SWEBench Pro (Public)55.462.1
Tau 3 Banking2626.8
TauBench V3 - Banking30.134.6
Terminal Bench 2.1 (Best Harness)6482.7
Terminal-Bench 2.172.182.7
Terminal-Bench Hard46.250.8
tool_decathlon52.848.2
toolathlon51.848.2
τ²-Bench Telecom (AA run)96.299.1
τ³-Bench25.826.8

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.