GLM-5.2
GLM-5.2 is capable in long context, coding, reasoning, and agentic tasks; and behind the leaders in factuality, instruction following, and math. Too few results yet to rate safety, multimodal tasks, or multilingual tasks.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Safety, Multimodal or Multilingual.
Price
$1.40input$4.40outputper million tokens
From Artificial Analysis · 8 providers tracked · All prices
Evidence
123results on73benchmarks
- 21 independently verified
- 34 aggregator
- 31 vendor-reported
- 37 cross-referenced
From 23 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingStructured outputsReasoning
As listed by OpenRouter
Research
35 papers reference GLM-5.2GLM-5.2 benchmark results
123 results on 73 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
14.0% behind the leader1 of 3 ranked benchmarks measured
- 78.33Oct 8, 2026
Show 1 more long context resultHide 1 long context result
- 42.33Oct 8, 2026
16.2% behind the leader7 of 10 ranked benchmarks measured
- LiveBench · Coding79.65Oct 8, 2026livebench_coding@2026-06-25
- 1602.95Sep 21, 2026
- Terminal-Bench 2.177.90Oct 8, 2026terminalbenchV21
- Terminal-Bench Hard50.76Oct 8, 2026aa_terminalbench_hard
- SciCode51.16Oct 8, 2026aa_scicode
- 62.10Oct 7, 2026
- LiveBench · Agentic Coding51.77Oct 8, 2026livebench_agentic_coding@2026-06-25
Show 7 more coding resultsHide 7 coding results
- SciCode36.11Sep 4, 2026aa_scicode
- 50.50Jul 27, 2026
- Terminal-Bench 2.151.69Oct 8, 2026terminalbenchV21
- Terminal-Bench 2.182.70Jul 13, 2026Terminal Bench 2.1 (Best Reported Harness)
- Terminal-Bench 2.181.00Jul 13, 2026Terminal Bench 2.1 (Terminus-2)
- 81.00Aug 13, 2026
- 82.70Jul 27, 2026
18.0% behind the leader6 of 6 ranked benchmarks measured
- GPQA Diamond89.49Oct 8, 2026gpqa
- LiveBench · Reasoning78.63Oct 8, 2026livebench_reasoning@2026-06-25
- 58.80Jun 18, 2026
- 20.86Oct 8, 2026
- Humanity's Last Exam41.15Oct 8, 2026aa_hle
- 22.78Jun 25, 2026
Show 12 more reasoning resultsHide 12 reasoning results
- 3.14Oct 8, 2026
- 16.70Oct 7, 2026
- 20.90Jun 16, 2026
- 20.90Jul 30, 2026
- GPQA Diamond68.59Oct 8, 2026gpqa
- GPQA Diamond91.20Oct 7, 2026GPQA
- 91.20Jul 27, 2026
- 89.50Jul 16, 2026
- Humanity's Last Exam9.78Oct 8, 2026aa_hle
- 54.70Oct 7, 2026
- 40.50Jul 13, 2026
- Humanity's Last Exam40.10Jul 16, 2026HLE (text only)
24.2% behind the leader4 of 7 ranked benchmarks measured
- 76.80Oct 7, 2026
- τ-Bench V3 · Banking34.64Oct 8, 2026tauBanking
- 43.70Oct 8, 2026
- Terminal-Bench 4.01.01Oct 8, 2026
Show 8 more agentic resultsHide 8 agentic results
- 33.70Oct 8, 2026
- AA ApexAgents35.60Jul 27, 2026APEX-Agents
- 42.66Oct 8, 2026
- 37.52Oct 8, 2026
- 77.80Oct 8, 2026
- 82.60Jul 27, 2026
- 77.80Jul 16, 2026
- τ-Bench V3 · Banking16.70Oct 8, 2026tauBanking
25.6% behind the leader3 of 4 ranked benchmarks measured
- AA-Omniscience · Non-hallucination73.70Oct 8, 2026omniscienceNonHallucination
- 34.20Jun 26, 2026
- AA-Omniscience · Accuracy24.33Oct 8, 2026omniscienceAccuracy
Show 3 more factuality resultsHide 3 factuality results
- AA-Omniscience · Accuracy20.28Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination66.32Oct 8, 2026omniscienceNonHallucination
- 38.10Jul 16, 2026
34.0% behind the leader2 of 3 ranked benchmarks measured
- IFBench73.33Oct 8, 2026aa_ifbench
- LiveBench · Instruction Following62.29Oct 8, 2026livebench_instruction_following@2026-06-25
Show 1 more instruction following resultHide 1 instruction following result
- 73.30Jul 16, 2026
43.7% behind the leader5 of 5 ranked benchmarks measured
- LiveBench · Mathematics89.78Oct 8, 2026livebench_math@2026-06-25
- 59.21Sep 9, 2026
- 29.27Sep 8, 2026
Show 6 more math resultsHide 6 math results
- 90.00Sep 2, 2026
- 99.20Oct 7, 2026
- 99.20Jul 16, 2026
- 92.42Sep 2, 2026
- HMMT Feb 202692.50Oct 7, 2026HMMT Feb 26
- 91.00Oct 7, 2026
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- livebench_language76.24Oct 8, 2026livebench_language@2026-06-25
- livebench_data_analysis73.74Oct 8, 2026livebench_data_analysis@2026-06-25
- 43.67Oct 8, 2026
- 1476Oct 5, 2026
- AA Intelligence22.00Jul 11, 2026Artificial Analysis Intelligence Index
- 77.00Jun 25, 2026
Show 53 more resultsHide 53 results
- 39.42Sep 9, 2026
- 35.12Sep 4, 2026
- AA Intelligence51.00Jun 18, 2026Artificial Analysis Intelligence Index
- AA Intelligence22.43Oct 8, 2026aa_intelligence_index
- AA Intelligence33.71Oct 8, 2026aa_intelligence_index
- 4.43Oct 8, 2026
- -6.57Oct 8, 2026
- 23.80Aug 13, 2026
- 20.40Jul 27, 2026
- 23.80Aug 28, 2026
- Artificial Analysis Coding Index68.76Sep 9, 2026aa_coding_index
- Artificial Analysis Coding Index46.49Sep 9, 2026aa_coding_index
- AutomationBench Public12.90Aug 13, 2026AutomationBench (Public)
- 77.20Aug 28, 2026
- 46.20Oct 7, 2026
- 46.20Aug 13, 2026
- 44.00Oct 7, 2026
- 61.80Aug 13, 2026
- 54.50Aug 13, 2026
- 1508.00Aug 28, 2026
- 1514.00Jul 30, 2026
- GDPval-AA v2 Elo1510.00Jul 27, 2026GDPval-AA v2 (Elo)
- 89.20Jul 16, 2026
- HLE (with tools)54.70Aug 28, 2026HLE w/ Tools
- 54.70Jul 16, 2026
- 54.70Aug 13, 2026
- 94.40Oct 7, 2026
- 94.40Jul 13, 2026
- 43.40Jul 27, 2026
- MLS Bench Litelower is better40.40Jul 27, 2026
- 48.90Oct 7, 2026
- 48.90Aug 13, 2026
- 41.40Jul 27, 2026
- 34.30Oct 7, 2026
- 31.70Aug 28, 2026
- 34.30Jul 27, 2026
- 63.70Oct 7, 2026
- 63.70Jul 27, 2026
- 9.50Aug 28, 2026
- 71.10Jul 27, 2026
- 28.10Jul 27, 2026
- 98.50Jul 16, 2026
- 13.00Oct 7, 2026
- SWE-Marathon19.40Aug 28, 2026SWE-Marathon (v1.1)
- 13.00Jul 27, 2026
- SWEBench Pro Public62.10Jul 16, 2026SWEBench Pro (Public)
- 4.60Aug 28, 2026
- 48.20Jul 13, 2026
- 48.20Oct 7, 2026
- 59.90Aug 28, 2026
- 59.90Aug 13, 2026
- τ²-Bench Telecom (AA run)99.12Oct 8, 2026aa_tau2
- τ³-Bench Banking26.80Jul 30, 2026τ³-Banking
GLM-5.2: common questions
Who makes GLM-5.2?
GLM-5.2 is made by Z.ai.
When was GLM-5.2 released?
GLM-5.2 was released on Jun 16, 2026, according to Artificial Analysis.
What is GLM-5.2 good at?
GLM-5.2 is capable in long context, coding, reasoning, and agentic tasks; and behind the leaders in factuality, instruction following, and math. Too few results yet to rate safety, multimodal tasks, or multilingual tasks.
How much does GLM-5.2 cost?
GLM-5.2 costs $1.40 per million input tokens and $4.40 per million output tokens, according to Artificial Analysis. We track its price at 8 providers. At a mix of three input tokens to one output token, it costs more than 73% of the 331 priced models we track.
How many benchmarks has GLM-5.2 been tested on?
We track 123 results for GLM-5.2 on 73 benchmarks from 23 sources, 21 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does GLM-5.2 support?
OpenRouter lists tool calling, structured outputs, and reasoning for GLM-5.2.
About this record
Where GLM-5.2's numbers come from, and every name it appears under.
- Tracked since
- Jun 15, 2026
- Newest source mention
- Aug 28, 2026
Where the results come from
Verification: 123 scores · 21 independently verified · 34 aggregator-attributed · 37 vendor cross-reference · 31 vendor-reported. How these tiers are assigned
From 23 sources on 12 sites. Hugging Face supplies 49 of them; the 21 independently verified results come from 9 sites. Bars are coloured by trust tier.
- huggingface.co49
- artificialanalysis.ai36
- api.llm-stats.com16
- livebench.ai7
- epoch.ai3
- thinkingmachines.ai3
- arcprize.org2
- labs.scale.com2
- matharena.ai2
- datasets-server.huggingface.co1
- lmarena.ai1
- simple-bench.com1