GLM-4.6
GLM-4.6 is behind the leaders in long context, reasoning, instruction following, agentic tasks, and coding. Too few results yet to rate safety, math, multimodal tasks, multilingual tasks, or factuality.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Safety, Math, Multimodal, Multilingual or Factuality.
Price
$0.57input$2.20outputper million tokens
From Artificial Analysis · 5 providers tracked · All prices
Evidence
87results on49benchmarks
- 11 independently verified
- 32 aggregator
- 18 vendor-reported
- 26 cross-referenced
From 13 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingStructured outputsReasoning
As listed by OpenRouter
GLM-4.6 benchmark results
87 results on 49 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
34.6% behind the leader1 of 3 ranked benchmarks measured
- 54.00Oct 8, 2026
Show 1 more long context resultHide 1 long context result
- 26.33Oct 8, 2026
37.2% behind the leader3 of 6 ranked benchmarks measured
- GPQA Diamond77.98Oct 8, 2026gpqa
- Humanity's Last Exam14.46Oct 8, 2026aa_hle
- 1.14Oct 8, 2026
Show 8 more reasoning resultsHide 8 reasoning results
- 0.00Oct 8, 2026
- GPQA Diamond63.23Oct 8, 2026gpqa
- GPQA Diamond81.00Oct 7, 2026GPQA
- 78.00Jun 6, 2026
- Humanity's Last Exam5.47Oct 8, 2026aa_hle
- 17.20Oct 7, 2026
- Humanity's Last Exam17.20Jun 25, 2026HLE_text
- Humanity's Last Exam13.30Jun 6, 2026HLE (w/o tools)
42.2% behind the leader2 of 3 ranked benchmarks measured
- 54.90Jun 25, 2026
- IFBench43.40Oct 8, 2026aa_ifbench
42.5% behind the leader3 of 7 ranked benchmarks measured
- 45.10Oct 7, 2026
- τ-Bench V3 · Banking13.40Oct 8, 2026tauBanking
- 13.04Oct 8, 2026
Show 2 more agentic resultsHide 2 agentic results
- 45.10Jun 6, 2026
- 24.29Jun 15, 2026
42.5% behind the leader8 of 10 ranked benchmarks measured
- 82.80Oct 7, 2026
- 68.00Oct 7, 2026
- SciCode38.43Sep 4, 2026aa_scicode
- 53.80May 18, 2026
- Terminal-Bench 2.149.44Oct 8, 2026terminalbenchV21
- 1340.60May 22, 2026
- Terminal-Bench Hard25.00Oct 8, 2026aa_terminalbench_hard
- 9.67Oct 8, 2026
Show 8 more coding resultsHide 8 coding results
- SciCode33.10Sep 4, 2026aa_scicode
- 38.00Jun 6, 2026
- 53.80Jun 6, 2026
- 55.40May 1, 2026
- 68.00Jun 15, 2026
- Terminal-Bench Hard28.79Oct 8, 2026aa_terminalbench_hard
- 23.60May 18, 2026
- 23.00Jun 6, 2026
0 of 5 ranked benchmarks measured
- 73.50May 18, 2026
0 of 4 ranked benchmarks measured
- 9.50May 2, 2026
- AA-Omniscience · Accuracy21.42Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination32.43Oct 8, 2026omniscienceNonHallucination
Show 2 more factuality resultsHide 2 factuality results
- AA-Omniscience · Accuracy26.88Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination5.95Oct 8, 2026omniscienceNonHallucination
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- frontiermath_tier_4_v12.13May 20, 2026frontiermath_tier_4
- 91.67May 20, 2026
- 91.67May 17, 2026
- vectara_avg_summary_length77.20May 2, 2026Average Summary Length (Words)
- vectara_answer_rate94.50May 2, 2026Answer Rate
- vectara_factual_consistency90.50May 2, 2026Factual Consistency Rate
Show 37 more resultsHide 37 results
- 18.62Sep 4, 2026
- 42.89Jun 18, 2026
- AA Intelligence14.93Oct 8, 2026aa_intelligence_index
- AA Intelligence18.53Oct 8, 2026aa_intelligence_index
- -31.68Oct 8, 2026
- -41.88Oct 8, 2026
- 93.90Oct 7, 2026
- AIME 202586.00Jun 6, 2026AIME25
- 59.80Jun 6, 2026
- Artificial Analysis Coding Index45.77Sep 9, 2026aa_coding_index
- 30.23Jun 18, 2026
- browsecomp_with_context_manager57.50May 18, 2026BrowseComp (w/ Context Manage)
- 49.50May 18, 2026
- 49.50Jun 6, 2026
- 29.20Jun 6, 2026
- 71.90Jun 6, 2026
- HLE (with tools)30.40May 18, 2026HLE (w/ Tools)
- HLE (with tools)30.40Jun 6, 2026HLE (w/ tools)
- 88.70Jun 25, 2026
- 89.20May 18, 2026
- 87.70May 18, 2026
- 79.50Jun 25, 2026
- LiveCodeBench70.00Jun 6, 2026LiveCodeBench (LCB)
- 83.20May 18, 2026
- 83.00Jun 6, 2026
- 30.00Jun 6, 2026
- 55.40May 1, 2026
- 40.50Sep 23, 2026
- 40.50Jun 6, 2026
- 24.50May 18, 2026
- Terminal-Bench 2.024.60Jun 15, 2026Terminal Bench 2
- theagentcompany35.00Jun 6, 2026AgentCompany
- 70.00Jun 6, 2026
- 75.20Jun 13, 2026
- τ²-Bench71.00Jun 6, 2026τ²-Bench-Telecom
- τ²-Bench Telecom (AA run)76.90Oct 8, 2026aa_tau2
- τ²-Bench Telecom (AA run)70.47Oct 8, 2026aa_tau2
GLM-4.6: common questions
Who makes GLM-4.6?
GLM-4.6 is made by Z.ai.
When was GLM-4.6 released?
GLM-4.6 was released on Sep 30, 2025, according to Artificial Analysis.
What is GLM-4.6 good at?
GLM-4.6 is behind the leaders in long context, reasoning, instruction following, agentic tasks, and coding. Too few results yet to rate safety, math, multimodal tasks, multilingual tasks, or factuality.
How much does GLM-4.6 cost?
GLM-4.6 costs $0.57 per million input tokens and $2.20 per million output tokens, according to Artificial Analysis. We track its price at 5 providers. At a mix of three input tokens to one output token, it costs more than 58% of the 330 priced models we track.
How many benchmarks has GLM-4.6 been tested on?
We track 87 results for GLM-4.6 on 49 benchmarks from 13 sources, 11 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does GLM-4.6 support?
OpenRouter lists tool calling, structured outputs, and reasoning for GLM-4.6.
About this record
Where GLM-4.6's numbers come from, and every name it appears under.
- Tracked since
- Apr 27, 2026
- Newest source mention
- Aug 23, 2026
Where the results come from
Verification: 87 scores · 11 independently verified · 32 aggregator-attributed · 26 vendor cross-reference · 18 vendor-reported. How these tiers are assigned
From 13 sources on 9 sites. Hugging Face supplies 37 of them; the 11 independently verified results come from 6 sites. Bars are coloured by trust tier.
- huggingface.co37
- artificialanalysis.ai32
- api.llm-stats.com7
- raw.githubusercontent.com4
- matharena.ai2
- swebench.com2
- datasets-server.huggingface.co1
- epoch.ai1
- labs.scale.com1