VECTOR WIREAI INTELLIGENCE
UTC
Refresh Models Deals Regulatory Sources

GLM-5.1 vs Hy4

14 SHARED BENCHMARKS

Across 14 shared benchmarks, GLM-5.1 scores higher on 1 and Hy4 on 13. The widest gap is critpt, where Hy4 scores 16.9 against 4.6. Hy4 is the cheaper of the two on tracked API pricing ($0.83 against $1.38 per million input tokens).

Z.AIVSTENCENT14 SHARED113 HEAD-TO-HEAD
BenchmarkGLM-5.1MARGINHy4
critpt4.616.9
cybergym68.778.4
DeepSWE1864.3
GPQA Diamond86.892.3
MCP Atlas75.683.7
nl2repo42.758.9
toolathlon40.774.1

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing. Quoted rates are the price-setter row we currently track for each model — its direct or vendor-official listing where one exists (direct, tencent), otherwise the lowest tracked offer.

Compare GLM-5.1 withEVERY TRACKED PAIRING

Anthropic12 PAIRINGS
OpenAI11 PAIRINGS
Google11 PAIRINGS
Alibaba11 PAIRINGS
DeepSeek5 PAIRINGS
Moonshot4 PAIRINGS
Meta2 PAIRINGS
Z.ai1 PAIRING
MiniMax1 PAIRING
NVIDIA1 PAIRING