VECTOR WIREAI INTELLIGENCE
NVDA$1,847+3.2%MSFT$512+1.1%GOOGL$199-0.4%META$728+2.7%AMD$184-1.2%TSM$212+0.6%PLTR$98+4.1%AI IDX4,821+1.9%
PKT
SEEDRefresh Models Deals Regulatory Sources

Claude Fable 5 vs GLM-5.1

AnthropicvsZ.ai36 shared benchmarks324 head-to-head
BenchmarkClaude Fable 5GLM-5.1
AA Agentic Index56.666
AA Intelligence62.141
AA-LCR76.768
AA-Omniscience43.30.8
arena_elo15081468
arena_text_factuality14861458
Artificial Analysis Coding Index76.555.8
browsecomp8879.3
coding_arena_elo16261510
critpt28.64.6
DeepSWE7018
Finance Agent v256.344.8
gdpval61.949.5
GDPval-AA (Elo)19321535
GPQA Diamond92.686.8
HLE64.552.3
HLE (with tools)64.552.3
IFBench63.576.3
livebench78.370.2
MCP Atlas84.775.6
OmniScience Accuracy65.325.2
OmniScience Non-Hallucination36.470.1
PostTrainBench41.420.1
Program Bench76.850.9
scicode60.243.8
simplebench81.955.1
simpleqa_verified68.338.1
SWE-bench Multilingual86.673.3
SWE-bench Pro8058.4
SWE-Marathon351
TauBench V3 - Banking38.113.6
Terminal-Bench 2.18863.5
Terminal-Bench Hard62.943.2
toolathlon61.740.7
τ²-Bench Telecom (AA run)98.597.7
τ³-Bench26.870.6

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.