GLM-5.3
GLM-5.3 is capable in agentic tasks, long context, reasoning, factuality, and coding; and behind the leaders in math and instruction following. Too few results yet to rate safety, multimodal tasks, or multilingual tasks.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Safety, Multimodal or Multilingual.
Price
$1.40input$4.40outputper million tokens
From Artificial Analysis · 8 providers tracked · All prices
Evidence
70results on49benchmarks
- 15 independently verified
- 26 aggregator
- 18 vendor-reported
- 11 cross-referenced
From 13 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingStructured outputsReasoning
As listed by OpenRouter
Research
14 papers reference GLM-5.3GLM-5.3 benchmark results
70 results on 49 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
10.5% behind the leader4 of 7 ranked benchmarks measured
- τ-Bench V3 · Banking50.31Oct 8, 2026tauBanking
- 84.20Oct 8, 2026
- Terminal-Bench 4.041.92Oct 8, 2026
- 57.54Oct 8, 2026
Show 3 more agentic resultsHide 3 agentic results
- 46.08Oct 8, 2026
- 41.24Oct 8, 2026
- Terminal-Bench 4.034.85Oct 8, 2026
11.7% behind the leader1 of 3 ranked benchmarks measured
- 79.67Oct 8, 2026
Show 1 more long context resultHide 1 long context result
- 72.67Oct 8, 2026
14.9% behind the leader5 of 6 ranked benchmarks measured
- GPQA Diamond91.72Oct 8, 2026gpqa
- LiveBench · Reasoning85.80Oct 8, 2026livebench_reasoning@2026-06-25
- 66.20Aug 21, 2026
- Humanity's Last Exam42.26Oct 8, 2026aa_hle
- 19.14Oct 8, 2026
Show 5 more reasoning resultsHide 5 reasoning results
- 14.57Oct 8, 2026
- GPQA Diamond88.10Sep 10, 2026GPQA Diamond (Pass@1)
- Humanity's Last Exam36.56Oct 8, 2026aa_hle
- 62.50Oct 7, 2026
- Humanity's Last Exam42.00Sep 10, 2026HLE (Pass@1)
18.5% behind the leader3 of 4 ranked benchmarks measured
- AA-Omniscience · Non-hallucination70.45Oct 8, 2026omniscienceNonHallucination
- 41.00Sep 21, 2026
- AA-Omniscience · Accuracy33.85Oct 8, 2026omniscienceAccuracy
Show 2 more factuality resultsHide 2 factuality results
- AA-Omniscience · Accuracy33.92Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination35.96Oct 8, 2026omniscienceNonHallucination
22.1% behind the leader5 of 10 ranked benchmarks measured
- Terminal-Bench 2.183.90Oct 8, 2026terminalbenchV21
- 1622.27Sep 21, 2026
- SciCode59.03Oct 8, 2026aa_scicode
- LiveBench · Coding78.95Oct 8, 2026livebench_coding@2026-06-25
- LiveBench · Agentic Coding60.91Oct 8, 2026livebench_agentic_coding@2026-06-25
Show 3 more coding resultsHide 3 coding results
- SciCode42.01Oct 8, 2026aa_scicode
- 88.20Oct 7, 2026
- Terminal-Bench 2.188.20Sep 10, 2026Terminal-Bench 2.1 (Pass@1)
35.2% behind the leader3 of 5 ranked benchmarks measured
- LiveBench · Mathematics87.90Oct 8, 2026livebench_math@2026-06-25
- 68.77Sep 21, 2026
- 29.27Sep 21, 2026
40.6% behind the leader1 of 3 ranked benchmarks measured
- LiveBench · Instruction Following69.30Oct 8, 2026livebench_instruction_following@2026-06-25
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- livebench_language79.86Oct 8, 2026livebench_language@2026-06-25
- livebench_data_analysis70.24Oct 8, 2026livebench_data_analysis@2026-06-25
- 1478Oct 6, 2026
- AA Intelligence45.00Sep 21, 2026Artificial Analysis Intelligence Index
- AA Intelligence34.30Oct 8, 2026aa_intelligence_index
- -8.40Oct 8, 2026
Show 28 more resultsHide 28 results
- 53.38Sep 9, 2026
- AA Intelligence44.78Oct 8, 2026aa_intelligence_index
- 14.30Oct 8, 2026
- 28.50Oct 7, 2026
- Agents' Last Exam28.50Sep 10, 2026Agent's Last Exam (Pass@1)
- 28.50Aug 28, 2026
- Artificial Analysis Coding Index74.76Sep 9, 2026aa_coding_index
- 84.50Oct 7, 2026
- CyberGym84.50Sep 10, 2026CyberGym (Pass@1)
- DeepSWE66.90Aug 28, 2026DeepSWE (v1.1)
- 66.90Oct 7, 2026
- DeepSWE v1.1 (Resolved)66.90Sep 10, 2026
- 1769.00Aug 28, 2026
- HLE (with tools)62.50Aug 28, 2026HLE w/ Tools
- HLE (with tools)62.50Sep 10, 2026HLE w/ tools (Pass@1)
- 58.00Oct 7, 2026
- NL2Repo-Bench (Score)58.00Sep 10, 2026
- 39.80Oct 7, 2026
- 19.00Oct 7, 2026
- 19.00Aug 28, 2026
- ProgramBench (Almost@1)19.00Sep 10, 2026
- 42.50Oct 7, 2026
- 28.30Oct 7, 2026
- Terminal-bench 3.028.30Sep 10, 2026Terminal-Bench 3.0 (Pass@1)
- Terminal-Bench 4.041.80Oct 7, 2026
- Terminal-Bench 4.037.90Sep 10, 2026Terminal-Bench 4.0 (Pass@1)
- 73.00Oct 7, 2026
- 73.00Aug 28, 2026
GLM-5.3: common questions
Who makes GLM-5.3?
GLM-5.3 is made by Z.ai.
When was GLM-5.3 released?
GLM-5.3 was released on Aug 18, 2026, according to Artificial Analysis.
What is GLM-5.3 good at?
GLM-5.3 is capable in agentic tasks, long context, reasoning, factuality, and coding; and behind the leaders in math and instruction following. Too few results yet to rate safety, multimodal tasks, or multilingual tasks.
How much does GLM-5.3 cost?
GLM-5.3 costs $1.40 per million input tokens and $4.40 per million output tokens, according to Artificial Analysis. We track its price at 8 providers. At a mix of three input tokens to one output token, it costs more than 73% of the 330 priced models we track.
How many benchmarks has GLM-5.3 been tested on?
We track 70 results for GLM-5.3 on 49 benchmarks from 13 sources, 15 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does GLM-5.3 support?
OpenRouter lists tool calling, structured outputs, and reasoning for GLM-5.3.
About this record
Where GLM-5.3's numbers come from, and every name it appears under.
- Tracked since
- Aug 14, 2026
- Newest source mention
- Oct 8, 2026
Where the results come from
Verification: 70 scores · 15 independently verified · 26 aggregator-attributed · 11 vendor cross-reference · 18 vendor-reported. How these tiers are assigned
From 13 sources on 9 sites. Artificial Analysis supplies 27 of them; the 15 independently verified results come from 7 sites. Bars are coloured by trust tier.
- artificialanalysis.ai27
- huggingface.co17
- api.llm-stats.com12
- livebench.ai7
- epoch.ai3
- datasets-server.huggingface.co1
- labs.scale.com1
- lmarena.ai1
- simple-bench.com1
Also known as
How our sources name GLM-5.3 at each reasoning setting.
| Setting | Short form | API id |
|---|---|---|
| low | — | glm-5-3-low glm-5.3 (low) |
| max | glm-5p3 max | glm-5.3 (max) glm-5.3-max |