GLM-4.5V
GLM-4.5V is behind the leaders in factuality, multimodal tasks, and agentic tasks. Too few results yet to rate reasoning, coding, safety, long context, math, multilingual tasks, or instruction following.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Reasoning, Coding, Safety, Long Context, Math, Multilingual or Instruction Following.
Price
$0.60input$1.80outputper million tokens
From Artificial Analysis · 4 providers tracked · All prices
Evidence
33results on18benchmarks
- 3 independently verified
- 30 aggregator
From 4 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingJSON modeReasoning
As listed by OpenRouter
Research
1 paper reference GLM-4.5VGLM-4.5V benchmark results
33 results on 18 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
38.3% behind the leader2 of 4 ranked benchmarks measured
- AA-Omniscience · Accuracy20.83Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination15.22Oct 8, 2026omniscienceNonHallucination
Show 2 more factuality resultsHide 2 factuality results
- AA-Omniscience · Accuracy18.22Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination9.76Oct 8, 2026omniscienceNonHallucination
39.7% behind the leader1 of 6 ranked benchmarks measured
- MMMU-Pro50.46Oct 8, 2026aa_mmmu_pro
Show 3 more multimodal resultsHide 3 multimodal results
- 1154.81Aug 25, 2026
- 1156May 1, 2026
- MMMU-Pro42.77Oct 8, 2026aa_mmmu_pro
43.5% behind the leader1 of 7 ranked benchmarks measured
- 0.49Jun 15, 2026
Show 1 more agentic resultHide 1 agentic result
- 0.00Jun 15, 2026
0 of 6 ranked benchmarks measured
- 0.00Oct 8, 2026
- GPQA Diamond57.27Oct 8, 2026gpqa
- GPQA Diamond68.38Oct 8, 2026gpqa
Show 2 more reasoning resultsHide 2 reasoning results
- Humanity's Last Exam3.48Oct 8, 2026aa_hle
- Humanity's Last Exam6.30Oct 8, 2026aa_hle
0 of 10 ranked benchmarks measured
- Terminal-Bench Hard6.82Oct 8, 2026aa_terminalbench_hard
- Terminal-Bench Hard5.30Oct 8, 2026aa_terminalbench_hard
- SciCode18.75Sep 4, 2026aa_scicode
Show 1 more coding resultHide 1 coding result
- SciCode22.11Sep 4, 2026aa_scicode
0 of 3 ranked benchmarks measured
- 0.00Oct 8, 2026
0 of 3 ranked benchmarks measured
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- 67.60Sep 2, 2026
- -55.58Oct 8, 2026
- τ²-Bench Telecom (AA run)19.59Oct 8, 2026aa_tau2
- AA Intelligence6.70Oct 8, 2026aa_intelligence_index
- τ²-Bench Telecom (AA run)22.51Oct 8, 2026aa_tau2
- AA Intelligence7.56Oct 8, 2026aa_intelligence_index
Show 5 more resultsHide 5 results
- 9.18Jun 18, 2026
- 10.91Jun 18, 2026
- -46.28Oct 8, 2026
- 10.80Jun 18, 2026
- Artificial Analysis Coding Index10.90Jun 18, 2026aa_coding_index
GLM-4.5V: common questions
Who makes GLM-4.5V?
GLM-4.5V is made by Z.ai.
When was GLM-4.5V released?
GLM-4.5V was released on Aug 11, 2025, according to Artificial Analysis.
What is GLM-4.5V good at?
GLM-4.5V is behind the leaders in factuality, multimodal tasks, and agentic tasks. Too few results yet to rate reasoning, coding, safety, long context, math, multilingual tasks, or instruction following.
How much does GLM-4.5V cost?
GLM-4.5V costs $0.60 per million input tokens and $1.80 per million output tokens, according to Artificial Analysis. We track its price at 4 providers. At a mix of three input tokens to one output token, it costs more than 56% of the 330 priced models we track.
How many benchmarks has GLM-4.5V been tested on?
We track 33 results for GLM-4.5V on 18 benchmarks from 4 sources, 3 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does GLM-4.5V support?
OpenRouter lists tool calling, json mode, and reasoning for GLM-4.5V.
About this record
Where GLM-4.5V's numbers come from, and every name it appears under.
- Tracked since
- Apr 27, 2026
- Newest source mention
- Aug 23, 2026
Where the results come from
Verification: 33 scores · 3 independently verified · 30 aggregator-attributed. How these tiers are assigned
From 4 sources on 4 sites. Artificial Analysis supplies 30 of them; the 3 independently verified results come from 3 sites. Bars are coloured by trust tier.
- artificialanalysis.ai30
- datasets-server.huggingface.co1
- lmarena.ai1
- matharena.ai1