GLM-4.7
GLM-4.7 is capable in long context; and behind the leaders in coding, instruction following, reasoning, and agentic tasks. Too few results yet to rate safety, math, multimodal tasks, multilingual tasks, or factuality.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Safety, Math, Multimodal, Multilingual or Factuality.
Price
$0.60input$2.20outputper million tokens
From Artificial Analysis · 6 providers tracked · All prices
Evidence
82results on49benchmarks
- 11 independently verified
- 32 aggregator
- 24 vendor-reported
- 15 cross-referenced
From 15 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingStructured outputsReasoning
As listed by OpenRouter
Research
8 papers reference GLM-4.7GLM-4.7 benchmark results
82 results on 49 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
24.4% behind the leader1 of 3 ranked benchmarks measured
- 71.00Oct 8, 2026
Show 1 more long context resultHide 1 long context result
- 40.67Oct 8, 2026
28.2% behind the leader7 of 10 ranked benchmarks measured
- 84.90Oct 7, 2026
- 73.80Oct 7, 2026
- 66.70Oct 7, 2026
- SciCode45.14Sep 4, 2026aa_scicode
- 1435.07May 22, 2026
- Terminal-Bench 2.145.32Oct 8, 2026terminalbenchV21
- Terminal-Bench Hard31.82Oct 8, 2026aa_terminalbench_hard
Show 5 more coding resultsHide 5 coding results
- 84.90May 30, 2026
- SciCode35.42Sep 4, 2026aa_scicode
- 73.80May 30, 2026
- Terminal-Bench Hard30.30Oct 8, 2026aa_terminalbench_hard
- 33.30May 18, 2026
30.2% behind the leader1 of 3 ranked benchmarks measured
- IFBench67.89Oct 8, 2026aa_ifbench
Show 1 more instruction following resultHide 1 instruction following result
- IFBench54.63Oct 8, 2026aa_ifbench
31.3% behind the leader4 of 6 ranked benchmarks measured
- GPQA Diamond85.86Oct 8, 2026gpqa
- 47.70May 10, 2026
- Humanity's Last Exam27.39Oct 8, 2026aa_hle
- 1.71Oct 8, 2026
Show 6 more reasoning resultsHide 6 reasoning results
- 0.00Oct 8, 2026
- GPQA Diamond66.36Oct 8, 2026gpqa
- GPQA Diamond85.70Oct 7, 2026GPQA
- Humanity's Last Exam6.39Oct 8, 2026aa_hle
- 42.80Oct 7, 2026
- 24.80May 18, 2026
43.7% behind the leader4 of 7 ranked benchmarks measured
- MCP Atlas52.00May 17, 2026MCP-Atlas (Public Set)
- 52.00Oct 7, 2026
- 25.83Oct 8, 2026
- τ-Bench V3 · Banking12.16Oct 8, 2026tauBanking
Show 3 more agentic resultsHide 3 agentic results
- 52.00May 30, 2026
- 33.69Jun 15, 2026
- 58.10Oct 8, 2026
0 of 5 ranked benchmarks measured
- 82.00Oct 7, 2026
- 82.00May 30, 2026
0 of 4 ranked benchmarks measured
- 32.20May 20, 2026
- 11.70May 2, 2026
- AA-Omniscience · Accuracy23.63Oct 8, 2026omniscienceAccuracy
Show 3 more factuality resultsHide 3 factuality results
- AA-Omniscience · Accuracy29.32Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination7.09Oct 8, 2026omniscienceNonHallucination
- AA-Omniscience · Non-hallucination6.98Oct 8, 2026omniscienceNonHallucination
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- 1442Jul 23, 2026
- frontiermath_tier_4_v10.00May 20, 2026frontiermath_tier_4
- 98.29May 20, 2026
- vectara_avg_summary_length70.60May 2, 2026Average Summary Length (Words)
- vectara_answer_rate99.80May 2, 2026Answer Rate
- vectara_factual_consistency88.30May 2, 2026Factual Consistency Rate
Show 35 more resultsHide 35 results
- 26.22Sep 4, 2026
- 54.34Jun 18, 2026
- AA Intelligence17.36Oct 8, 2026aa_intelligence_index
- AA Intelligence22.24Oct 8, 2026aa_intelligence_index
- -47.32Oct 8, 2026
- -36.43Oct 8, 2026
- 95.70Oct 7, 2026
- 95.70May 30, 2026
- 92.90May 17, 2026
- Artificial Analysis Coding Index45.26Sep 9, 2026aa_coding_index
- 32.01Jun 18, 2026
- browsecomp_with_context_manager67.50May 18, 2026BrowseComp (w/ Context Manage)
- browsecomp_with_context_manager67.50May 30, 2026BrowseComp (w/ Context Manager)
- 66.60Oct 7, 2026
- 66.60May 30, 2026
- 23.50Jun 13, 2026
- 61.90May 30, 2026
- HLE (with tools)42.80May 18, 2026HLE (w/ Tools)
- 97.10May 18, 2026
- HMMT Feb. 202597.10May 30, 2026HMMT 2025 (Feb.)
- 93.50May 18, 2026
- HMMT Nov. 202593.50May 30, 2026HMMT 2025 (Nov.)
- 84.30Oct 7, 2026
- 62.00May 30, 2026
- 33.30Sep 23, 2026
- 41.00Oct 7, 2026
- Terminal-Bench 2.032.80May 17, 2026Terminal-Bench 2.0 (Claude Code)
- 41.00May 30, 2026
- 23.80May 17, 2026
- xbench-DeepSearch52.30May 30, 2026xbench-DeepSearch (2025.10)
- 52.30May 3, 2026
- 87.40Jun 13, 2026
- 87.40Jun 12, 2026
- τ²-Bench Telecom (AA run)94.15Oct 8, 2026aa_tau2
- τ²-Bench Telecom (AA run)95.91Oct 8, 2026aa_tau2
GLM-4.7: common questions
Who makes GLM-4.7?
GLM-4.7 is made by Z.ai.
When was GLM-4.7 released?
GLM-4.7 was released on Dec 22, 2025, according to Artificial Analysis.
What is GLM-4.7 good at?
GLM-4.7 is capable in long context; and behind the leaders in coding, instruction following, reasoning, and agentic tasks. Too few results yet to rate safety, math, multimodal tasks, multilingual tasks, or factuality.
How much does GLM-4.7 cost?
GLM-4.7 costs $0.60 per million input tokens and $2.20 per million output tokens, according to Artificial Analysis. We track its price at 6 providers. At a mix of three input tokens to one output token, it costs more than 58% of the 331 priced models we track.
How many benchmarks has GLM-4.7 been tested on?
We track 82 results for GLM-4.7 on 49 benchmarks from 15 sources, 11 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does GLM-4.7 support?
OpenRouter lists tool calling, structured outputs, and reasoning for GLM-4.7.
About this record
Where GLM-4.7's numbers come from, and every name it appears under.
- Tracked since
- Apr 27, 2026
- Newest source mention
- Jun 3, 2026
Where the results come from
Verification: 82 scores · 11 independently verified · 32 aggregator-attributed · 15 vendor cross-reference · 24 vendor-reported. How these tiers are assigned
From 15 sources on 10 sites. Artificial Analysis supplies 32 of them; the 11 independently verified results come from 7 sites. Bars are coloured by trust tier.
- artificialanalysis.ai32
- huggingface.co27
- api.llm-stats.com12
- raw.githubusercontent.com4
- epoch.ai2
- datasets-server.huggingface.co1
- labs.scale.com1
- lmarena.ai1
- matharena.ai1
- simple-bench.com1