GLM-4.5
GLM-4.5 is behind the leaders in factuality, long context, coding, and instruction following. Too few results yet to rate reasoning, agentic tasks, safety, math, multimodal tasks, or multilingual tasks.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Reasoning, Agentic, Safety, Math, Multimodal or Multilingual.
Price
$0.60input$2.20outputper million tokens
From Z.ai's own price page · 2 providers tracked · All prices
Evidence
36results on29benchmarks
- 4 independently verified
- 15 aggregator
- 12 vendor-reported
- 5 cross-referenced
From 5 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingJSON modeReasoning
As listed by OpenRouter
Research
5 papers reference GLM-4.5GLM-4.5 benchmark results
36 results on 29 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
29.7% behind the leader2 of 4 ranked benchmarks measured
- AA-Omniscience · Accuracy25.13Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination29.85Oct 8, 2026omniscienceNonHallucination
35.6% behind the leader1 of 3 ranked benchmarks measured
- 52.67Oct 8, 2026
41.0% behind the leader4 of 10 ranked benchmarks measured
- 64.20Oct 7, 2026
- 52.70May 15, 2026
- SciCode34.84Sep 4, 2026aa_scicode
- Terminal-Bench Hard21.97Oct 8, 2026aa_terminalbench_hard
Show 4 more coding resultsHide 4 coding results
- 41.70Oct 7, 2026
- 54.20May 1, 2026
- 64.20May 1, 2026
- 64.20May 15, 2026
41.9% behind the leader1 of 3 ranked benchmarks measured
- IFBench44.08Oct 8, 2026aa_ifbench
0 of 6 ranked benchmarks measured
- 0.00Oct 8, 2026
- GPQA Diamond78.18Oct 8, 2026gpqa
- Humanity's Last Exam12.97Oct 8, 2026aa_hle
Show 2 more reasoning resultsHide 2 reasoning results
- GPQA Diamond79.10Oct 7, 2026GPQA
- 14.40Oct 7, 2026
0 of 7 ranked benchmarks measured
- 0.00Jun 15, 2026
- 26.40Oct 7, 2026
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- 93.33May 20, 2026
- 54.20May 1, 2026
- AA Intelligence12.77Oct 8, 2026aa_intelligence_index
- -27.38Oct 8, 2026
- τ²-Bench Telecom (AA run)42.98Oct 8, 2026aa_tau2
- 16.20Jun 18, 2026
Show 11 more resultsHide 11 results
- Artificial Analysis Coding Index26.26Jun 18, 2026aa_coding_index
- 77.80Oct 7, 2026
- 72.90Aug 23, 2026
- MATH-500 (EM)98.20Oct 7, 2026MATH-500
- 84.60Oct 7, 2026
- 31.70May 15, 2026
- 63.20May 15, 2026
- TAU-bench (airline)60.40Oct 7, 2026TAU-bench Airline
- TAU-bench (retail)79.70Oct 7, 2026TAU-bench Retail
- 37.50Sep 23, 2026
- 39.90May 15, 2026
GLM-4.5: common questions
Who makes GLM-4.5?
GLM-4.5 is made by Z.ai.
When was GLM-4.5 released?
GLM-4.5's weights were first published on Hugging Face on Jul 20, 2025.
What is GLM-4.5 good at?
GLM-4.5 is behind the leaders in factuality, long context, coding, and instruction following. Too few results yet to rate reasoning, agentic tasks, safety, math, multimodal tasks, or multilingual tasks.
How much does GLM-4.5 cost?
GLM-4.5 costs $0.60 per million input tokens and $2.20 per million output tokens, according to Z.ai's own price page. We track its price at 2 providers. At a mix of three input tokens to one output token, it costs more than 58% of the 331 priced models we track.
How many benchmarks has GLM-4.5 been tested on?
We track 36 results for GLM-4.5 on 29 benchmarks from 5 sources, 4 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does GLM-4.5 support?
OpenRouter lists tool calling, json mode, and reasoning for GLM-4.5.
About this record
Where GLM-4.5's numbers come from, and every name it appears under.
- Tracked since
- Apr 27, 2026
- Newest source mention
- Aug 24, 2026
Where the results come from
Verification: 36 scores · 4 independently verified · 15 aggregator-attributed · 5 vendor cross-reference · 12 vendor-reported. How these tiers are assigned
From 5 sources on 5 sites. Artificial Analysis supplies 15 of them; the 4 independently verified results come from 2 sites. Bars are coloured by trust tier.
- artificialanalysis.ai15
- api.llm-stats.com12
- huggingface.co5
- swebench.com3
- matharena.ai1