VECTOR WIREAI INTELLIGENCE
NVDA$1,847+3.2%MSFT$512+1.1%GOOGL$199-0.4%META$728+2.7%AMD$184-1.2%TSM$212+0.6%PLTR$98+4.1%AI IDX4,821+1.9%
PKT
SEEDRefresh Models Deals Regulatory Sources

GLM-5.2 Full Open Source vs GPT-5.6 Sol

Z.aivsOpenAI59 shared benchmarks553 head-to-head
BenchmarkGLM-5.2 Full Open SourceGPT-5.6 Sol
AA Agentic Index45.757.8
AA Intelligence52.661
AA-LCR76.777.7
AA-Omniscience4.422
Agents' Last Exam23.852.7
aime_202699.299.9
apex_agents35.656.7
arc_agi_17797.5
ARC-AGI-222.892.5
arena_elo14711482
arena_text_factuality14601474
Artificial Analysis Coding Index68.878.3
coding_arena_elo15821619
critpt20.932.3
DeepSWE46.273
DeepSWE 1.14473
FORTRESS (Adversarial)71.382.4
FORTRESS (Benign)9098.1
gdpval50.361.1
GDPval-AA v215141748
GDPval-AA v2 (Elo)15101736
Global-MMLU-Lite89.291.8
GPQA Diamond91.294.6
HiL-Bench43.732.3
HLE54.749.5
HLE (with tools)54.758
IFBench73.372.7
itbenchSre42.756.2
JobBench43.445.4
livebench_agentic_coding51.856.2
livebench_coding79.783.9
livebench_data_analysis73.779.8
livebench_instruction_following62.371.8
livebench_language76.287.7
livebench_math89.896.2
livebench_reasoning78.691.7
MCP Atlas82.683.6
MLS Bench Lite40.446.2
officeqa_pro41.463.2
OmniScience Accuracy24.359.4
OmniScience Non-Hallucination73.710.6
PostTrainBench34.334.6
Program Bench63.777.6
ResearchRubrics71.173.8
scicode50.556.9
simplebench58.864.8
simpleqa_verified38.171.6
SpreadsheetBench 228.132.4
strongreject98.598.5
SWE-bench Pro62.164.6
SWE-Marathon1339
SWEBench Pro (Public)62.164.6
Tau 3 Banking26.833
TauBench V3 - Banking34.644.3
Terminal Bench 2.1 (Best Harness)82.789.5
Terminal-Bench 2.182.789.5
Terminal-Bench Hard50.865.9
toolathlon48.258
τ²-Bench Telecom (AA run)99.185.1

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.