GLM-5.1 vs Qwen3.5 2B
20 SHARED BENCHMARKSAcross 20 shared benchmarks, GLM-5.1 scores higher on 20 and Qwen3.5 2B on 0. The widest gap is scicode, where GLM-5.1 scores 44.8 against 7.2.
Z.AIVSALIBABA20 SHARED20–0 HEAD-TO-HEAD
Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.
Compare GLM-5.1 withALL PAIRINGS →
Z.aiGLM-5.2