Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

GLM-4.6

Score basis

GLM-4.6 is behind the leaders in long context, reasoning, instruction following, agentic tasks, and coding. Too few results yet to rate safety, math, multimodal tasks, multilingual tasks, or factuality.

Capability profile

Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.

Limited5 capabilities
  1. Long Context−34.6%1 of 3
  2. Reasoning−37.2%3 of 6
  3. Instruction Following−42.2%2 of 3
  4. Agentic−42.5%3 of 7
  5. Coding−42.5%8 of 10
Not rated5 capabilities

Too few results yet to rate Safety, Math, Multimodal, Multilingual or Factuality.

Price

$0.57input$2.20outputper million tokens

From Artificial Analysis · 5 providers tracked · All prices

Costs more than 58% of 330 priced models · 3:1 input-to-output blend, log scale

Evidence

87results on49benchmarks

  • 11 independently verified
  • 32 aggregator
  • 18 vendor-reported
  • 26 cross-referenced

From 13 sources · latest Oct 8, 2026 · How verification works

API features

Tool callingStructured outputsReasoning

As listed by OpenRouter

GLM-4.6 benchmark results

87 results on 49 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.

Long Context

LimitedFull long context ranking

34.6% behind the leader1 of 3 ranked benchmarks measured

Show 1 more long context resultHide 1 long context result

Reasoning

LimitedFull reasoning ranking

37.2% behind the leader3 of 6 ranked benchmarks measured

Show 8 more reasoning resultsHide 8 reasoning results

Instruction Following

LimitedFull instruction following ranking

42.2% behind the leader2 of 3 ranked benchmarks measured

Show 2 more instruction following resultsHide 2 instruction following results

Agentic

LimitedFull agentic ranking

42.5% behind the leader3 of 7 ranked benchmarks measured

Show 2 more agentic resultsHide 2 agentic results

Coding

LimitedFull coding ranking

42.5% behind the leader8 of 10 ranked benchmarks measured

Show 8 more coding resultsHide 8 coding results

Math

Not enough dataFull math ranking

0 of 5 ranked benchmarks measured

Factuality

Not enough dataFull factuality ranking

0 of 4 ranked benchmarks measured

Show 2 more factuality resultsHide 2 factuality results

More results

Benchmarks outside the capability baskets. They are not ranked against a leader.

Show 37 more resultsHide 37 results

GLM-4.6: common questions

Who makes GLM-4.6?

GLM-4.6 is made by Z.ai.

When was GLM-4.6 released?

GLM-4.6 was released on Sep 30, 2025, according to Artificial Analysis.

What is GLM-4.6 good at?

GLM-4.6 is behind the leaders in long context, reasoning, instruction following, agentic tasks, and coding. Too few results yet to rate safety, math, multimodal tasks, multilingual tasks, or factuality.

How much does GLM-4.6 cost?

GLM-4.6 costs $0.57 per million input tokens and $2.20 per million output tokens, according to Artificial Analysis. We track its price at 5 providers. At a mix of three input tokens to one output token, it costs more than 58% of the 330 priced models we track.

How many benchmarks has GLM-4.6 been tested on?

We track 87 results for GLM-4.6 on 49 benchmarks from 13 sources, 11 of them independently verified. The latest was recorded on Oct 8, 2026.

Which API features does GLM-4.6 support?

OpenRouter lists tool calling, structured outputs, and reasoning for GLM-4.6.

About this record

Where GLM-4.6's numbers come from, and every name it appears under.

Tracked since
Apr 27, 2026
Newest source mention
Aug 23, 2026

Where the results come from

Verification: 87 scores · 11 independently verified · 32 aggregator-attributed · 26 vendor cross-reference · 18 vendor-reported. How these tiers are assigned

From 13 sources on 9 sites. Hugging Face supplies 37 of them; the 11 independently verified results come from 6 sites. Bars are coloured by trust tier.

  • huggingface.co37
  • artificialanalysis.ai32
  • api.llm-stats.com7
  • raw.githubusercontent.com4
  • matharena.ai2
  • swebench.com2
  • datasets-server.huggingface.co1
  • epoch.ai1
  • labs.scale.com1

Also known as

glm-4.6 (Non-Reasoning)glm-4.6 (reasoning)GLM-4.6 (T=1)glm-4-6-reasoningglm-4-6glm-4.6 (together)