Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

GLM-4.6V

Score basis

GLM-4.6V is behind the leaders in factuality, multimodal tasks, and agentic tasks. Too few results yet to rate reasoning, coding, safety, long context, math, multilingual tasks, or instruction following.

Capability profile

Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.

Limited3 capabilities
  1. Factuality−35.0%2 of 4
  2. Multimodal−40.1%1 of 6
  3. Agentic−43.0%1 of 7
Not rated7 capabilities

Too few results yet to rate Reasoning, Coding, Safety, Long Context, Math, Multilingual or Instruction Following.

Price

$0.30input$0.90outputper million tokens

From Artificial Analysis · 3 providers tracked · All prices

Cheaper than 62% of 330 priced models · 3:1 input-to-output blend, log scale

Evidence

33results on17benchmarks

  • 2 independently verified
  • 31 aggregator

From 3 sources · latest Oct 8, 2026 · How verification works

API features

Tool callingJSON modeReasoning

As listed by OpenRouter

GLM-4.6V benchmark results

33 results on 17 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.

Factuality

LimitedFull factuality ranking

35.0% behind the leader2 of 4 ranked benchmarks measured

Show 2 more factuality resultsHide 2 factuality results

Multimodal

LimitedFull multimodal ranking

40.1% behind the leader1 of 6 ranked benchmarks measured

Show 3 more multimodal resultsHide 3 multimodal results

Agentic

LimitedFull agentic ranking

43.0% behind the leader1 of 7 ranked benchmarks measured

Show 1 more agentic resultHide 1 agentic result

Reasoning

Not enough dataFull reasoning ranking

0 of 6 ranked benchmarks measured

Show 2 more reasoning resultsHide 2 reasoning results

Coding

Not enough dataFull coding ranking

0 of 10 ranked benchmarks measured

Show 1 more coding resultHide 1 coding result

Long Context

Not enough dataFull long context ranking

0 of 3 ranked benchmarks measured

Instruction Following

Not enough dataFull instruction following ranking

0 of 3 ranked benchmarks measured

More results

Benchmarks outside the capability baskets. They are not ranked against a leader.

Show 4 more resultsHide 4 results

GLM-4.6V: common questions

Who makes GLM-4.6V?

GLM-4.6V is made by Z.ai.

When was GLM-4.6V released?

GLM-4.6V was released on Dec 8, 2025, according to Artificial Analysis.

What is GLM-4.6V good at?

GLM-4.6V is behind the leaders in factuality, multimodal tasks, and agentic tasks. Too few results yet to rate reasoning, coding, safety, long context, math, multilingual tasks, or instruction following.

How much does GLM-4.6V cost?

GLM-4.6V costs $0.30 per million input tokens and $0.90 per million output tokens, according to Artificial Analysis. We track its price at 3 providers. At a mix of three input tokens to one output token, it is cheaper than 62% of the 330 priced models we track.

How many benchmarks has GLM-4.6V been tested on?

We track 33 results for GLM-4.6V on 17 benchmarks from 3 sources, 2 of them independently verified. The latest was recorded on Oct 8, 2026.

Which API features does GLM-4.6V support?

OpenRouter lists tool calling, json mode, and reasoning for GLM-4.6V.

About this record

Where GLM-4.6V's numbers come from, and every name it appears under.

Tracked since
Apr 25, 2026
Newest source mention
May 2, 2026

Where the results come from

Verification: 33 scores · 2 independently verified · 31 aggregator-attributed. How these tiers are assigned

From 3 sources on 3 sites. Artificial Analysis supplies 31 of them; the 2 independently verified results come from 2 sites. Bars are coloured by trust tier.

  • artificialanalysis.ai31
  • datasets-server.huggingface.co1
  • lmarena.ai1

Also known as

glm-4.6v (non-reasoning)glm-4-6v (Non-Reasoning)glm-4-6vglm-4.6v (reasoning)glm-4.6 (reasoning)glm-4-6v (reasoning)glm-4-6v-reasoning