Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

GLM-4.5-Base vs Qwen 2.5 14B

Z.ai · released

Alibaba · released

Scores updated · 7 tests both models report · How we compare

Every test, side by side

All 7 tests both models report. The winning score is in its model's colour; marks a score checked independently.

Other results7 tests
  • SimpleQAGLM-4.5-Base by 23.929.35.4+23.9
  • MATHQwen 2.5 14B by 14.66175.6+14.6
  • GPQA (unspecified)Qwen 2.5 14B by 9.433.542.9+9.4
  • MMLUGLM-4.5-Base by 7.887.779.9+7.8
  • HumanEvalGLM-4.5-Base by 6.178.272.1+6.1
  • DROPQwen 2.5 14B by 2.682.985.5+2.6
  • MGSMGLM-4.5-Base by 1.781.379.6+1.7

Questions people ask

How do you compare the two?

We use the 7 benchmark tests both models have published scores on. None of them is in the eight capability areas we count, so this page lists them without a verdict. Each score is the one shown on the model's own page.