GLM-4.5-Air
GLM-4.5-Air is behind the leaders in coding and agentic tasks. Too few results yet to rate reasoning, safety, long context, math, multimodal tasks, multilingual tasks, instruction following, or factuality.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Reasoning, Safety, Long Context, Math, Multimodal, Multilingual, Instruction Following or Factuality.
Price
$0.17input$0.98outputper million tokens
From Artificial Analysis · 4 providers tracked · All prices
Evidence
38results on35benchmarks
- 11 independently verified
- 15 aggregator
- 12 vendor-reported
From 10 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingReasoning
As listed by OpenRouter
Research
4 papers reference GLM-4.5-AirGLM-4.5-Air benchmark results
38 results on 35 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
41.0% behind the leader3 of 10 ranked benchmarks measured
- 57.60Oct 7, 2026
- SciCode30.56Sep 4, 2026aa_scicode
- Terminal-Bench Hard20.45Oct 8, 2026aa_terminalbench_hard
Show 1 more coding resultHide 1 coding result
- 37.30Oct 7, 2026
46.9% behind the leader2 of 7 ranked benchmarks measured
- 21.30Oct 7, 2026
- 3.02Jun 15, 2026
0 of 6 ranked benchmarks measured
- 0.00Oct 8, 2026
- GPQA Diamond73.33Oct 8, 2026gpqa
- Humanity's Last Exam7.04Oct 8, 2026aa_hle
Show 2 more reasoning resultsHide 2 reasoning results
- GPQA Diamond75.00Oct 7, 2026GPQA
- 10.60Oct 7, 2026
0 of 3 ranked benchmarks measured
- 46.67Oct 8, 2026
0 of 3 ranked benchmarks measured
- IFBench37.55Oct 8, 2026aa_ifbench
0 of 4 ranked benchmarks measured
- 9.30May 2, 2026
- AA-Omniscience · Accuracy16.25Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination7.12Oct 8, 2026omniscienceNonHallucination
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- 83.33May 20, 2026
- 57.09May 19, 2026
- 98.56May 10, 2026
- 99.00May 10, 2026
- 56.15May 10, 2026
- 97.80May 10, 2026
Show 16 more resultsHide 16 results
- 21.01Jun 18, 2026
- AA Intelligence11.09Oct 8, 2026aa_intelligence_index
- -61.53Oct 8, 2026
- 98.99May 10, 2026
- 23.82Jun 18, 2026
- 76.40Oct 7, 2026
- 70.70Aug 23, 2026
- MATH-500 (EM)98.10Oct 7, 2026MATH-500
- 81.40Oct 7, 2026
- TAU-bench (airline)60.80Oct 7, 2026TAU-bench Airline
- TAU-bench (retail)77.90Oct 7, 2026TAU-bench Retail
- 30.00Sep 23, 2026
- vectara_answer_rate98.10May 2, 2026Answer Rate
- vectara_avg_summary_length70.60May 2, 2026Average Summary Length (Words)
- vectara_factual_consistency90.70May 2, 2026Factual Consistency Rate
- τ²-Bench Telecom (AA run)46.49Oct 8, 2026aa_tau2
GLM-4.5-Air: common questions
Who makes GLM-4.5-Air?
GLM-4.5-Air is made by Z.ai.
When was GLM-4.5-Air released?
GLM-4.5-Air was released on Jul 28, 2025, according to Artificial Analysis.
What is GLM-4.5-Air good at?
GLM-4.5-Air is behind the leaders in coding and agentic tasks. Too few results yet to rate reasoning, safety, long context, math, multimodal tasks, multilingual tasks, instruction following, or factuality.
How much does GLM-4.5-Air cost?
GLM-4.5-Air costs $0.17 per million input tokens and $0.98 per million output tokens, according to Artificial Analysis. We track its price at 4 providers. At a mix of three input tokens to one output token, it is cheaper than 67% of the 330 priced models we track.
How many benchmarks has GLM-4.5-Air been tested on?
We track 38 results for GLM-4.5-Air on 35 benchmarks from 10 sources, 11 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does GLM-4.5-Air support?
OpenRouter lists tool calling and reasoning for GLM-4.5-Air.
About this record
Where GLM-4.5-Air's numbers come from, and every name it appears under.
- Tracked since
- Apr 27, 2026
- Newest source mention
- Aug 24, 2026
Where the results come from
Verification: 38 scores · 11 independently verified · 15 aggregator-attributed · 12 vendor-reported. How these tiers are assigned
From 10 sources on 5 sites. Artificial Analysis supplies 15 of them; the 11 independently verified results come from 3 sites. Bars are coloured by trust tier.
- artificialanalysis.ai15
- api.llm-stats.com12
- storage.googleapis.com6
- raw.githubusercontent.com4
- matharena.ai1