Grok 4.7
Grok 4.7 is strong in factuality; capable in long context and reasoning; and behind the leaders in multimodal tasks, agentic tasks, coding, instruction following, and math. Too few results yet to rate safety or multilingual tasks.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Safety or Multilingual.
Price
$2.00input$6.00outputper million tokens
From Artificial Analysis · 3 providers tracked · All prices
Evidence
65results on32benchmarks
- 27 independently verified
- 32 aggregator
- 6 vendor-reported
From 10 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingStructured outputsReasoning
As listed by OpenRouter
Grok 4.7 benchmark results
65 results on 32 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
5.6% behind the leader3 of 4 ranked benchmarks measured
- AA-Omniscience · Non-hallucination70.66Oct 8, 2026omniscienceNonHallucination
- 56.00Oct 2, 2026
- AA-Omniscience · Accuracy47.45Oct 8, 2026omniscienceAccuracy
Show 4 more factuality resultsHide 4 factuality results
- AA-Omniscience · Accuracy47.82Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Accuracy49.70Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination67.58Oct 8, 2026omniscienceNonHallucination
- AA-Omniscience · Non-hallucination60.07Oct 8, 2026omniscienceNonHallucination
16.7% behind the leader1 of 3 ranked benchmarks measured
- 76.67Oct 8, 2026
20.3% behind the leader4 of 6 ranked benchmarks measured
- LiveBench · Reasoning82.65Oct 8, 2026livebench_reasoning@2026-06-25
- Humanity's Last Exam43.14Oct 8, 2026aa_hle
- 61.39Oct 7, 2026
- 17.71Oct 8, 2026
Show 15 more reasoning resultsHide 15 reasoning results
- 16.67Oct 7, 2026
- 58.75Oct 7, 2026
- 58.33Oct 7, 2026
- 0.77Oct 7, 2026
- 8.37Oct 7, 2026
- 10.05Oct 7, 2026
- 7.02Oct 7, 2026
- 0.25Oct 7, 2026
- 1.71Oct 7, 2026
- 1.73Oct 7, 2026
- 1.82Oct 7, 2026
- 18.00Oct 8, 2026
- 13.43Oct 8, 2026
- Humanity's Last Exam39.16Oct 8, 2026aa_hle
- Humanity's Last Exam42.31Oct 8, 2026aa_hle
26.4% behind the leader1 of 6 ranked benchmarks measured
- MMMU-Pro78.15Oct 8, 2026aa_mmmu_pro
27.4% behind the leader2 of 7 ranked benchmarks measured
- 60.62Oct 8, 2026
- Terminal-Bench 4.025.76Oct 8, 2026
Show 5 more agentic resultsHide 5 agentic results
- 42.08Oct 8, 2026
- 54.43Oct 8, 2026
- 60.49Oct 8, 2026
- Terminal-Bench 4.024.75Oct 8, 2026
- Terminal-Bench 4.016.16Oct 8, 2026
31.3% behind the leader4 of 10 ranked benchmarks measured
- 1638.05Sep 23, 2026
- SciCode57.41Oct 8, 2026aa_scicode
- LiveBench · Coding77.16Oct 8, 2026livebench_coding@2026-06-25
- LiveBench · Agentic Coding53.99Oct 8, 2026livebench_agentic_coding@2026-06-25
31.3% behind the leader1 of 3 ranked benchmarks measured
- LiveBench · Instruction Following75.28Oct 8, 2026livebench_instruction_following@2026-06-25
38.2% behind the leader5 of 5 ranked benchmarks measured
- LiveBench · Mathematics95.68Oct 8, 2026livebench_math@2026-06-25
- 52.98Oct 3, 2026
- 17.07Oct 3, 2026
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- livebench_data_analysis76.88Oct 8, 2026livebench_data_analysis@2026-06-25
- livebench_language80.14Oct 8, 2026livebench_language@2026-06-25
- 63.33Oct 7, 2026
- 90.00Oct 7, 2026
- 90.17Oct 7, 2026
- 89.50Oct 7, 2026
Show 12 more resultsHide 12 results
- AA Intelligence46.45Oct 8, 2026aa_intelligence_index
- AA Intelligence46.33Oct 8, 2026aa_intelligence_index
- AA Intelligence42.22Oct 8, 2026aa_intelligence_index
- 32.03Oct 8, 2026
- 29.62Oct 8, 2026
- 30.90Oct 8, 2026
- 80.30Oct 7, 2026
- 71.00Oct 7, 2026
- 56.70Oct 7, 2026
- 46.00Oct 7, 2026
- Terminal-Bench 4.038.00Oct 7, 2026
- WMDP88.10Oct 7, 2026WMDP-Cyber
Grok 4.7: common questions
Who makes Grok 4.7?
Grok 4.7 is made by SpaceXAI.
When was Grok 4.7 released?
Grok 4.7 was released on Sep 21, 2026, according to SpaceXAI's own announcement.
What is Grok 4.7 good at?
Grok 4.7 is strong in factuality; capable in long context and reasoning; and behind the leaders in multimodal tasks, agentic tasks, coding, instruction following, and math. Too few results yet to rate safety or multilingual tasks.
How much does Grok 4.7 cost?
Grok 4.7 costs $2.00 per million input tokens and $6.00 per million output tokens, according to Artificial Analysis. We track its price at 3 providers. At a mix of three input tokens to one output token, it costs more than 75% of the 331 priced models we track.
How many benchmarks has Grok 4.7 been tested on?
We track 65 results for Grok 4.7 on 32 benchmarks from 10 sources, 27 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does Grok 4.7 support?
OpenRouter lists tool calling, structured outputs, and reasoning for Grok 4.7.
About this record
Where Grok 4.7's numbers come from, and every name it appears under.
- Tracked since
- Aug 6, 2026
- Newest source mention
- Oct 7, 2026
Where the results come from
Verification: 65 scores · 27 independently verified · 32 aggregator-attributed · 6 vendor-reported. How these tiers are assigned
From 10 sources on 6 sites. Artificial Analysis supplies 32 of them; the 27 independently verified results come from 4 sites. Bars are coloured by trust tier.
- artificialanalysis.ai32
- arcprize.org16
- livebench.ai7
- api.llm-stats.com6
- epoch.ai3
- datasets-server.huggingface.co1
Also known as
How our sources name Grok 4.7 at each reasoning setting.
| Setting | Short form | API id |
|---|---|---|
| low | grok 4.7 (low) | grok-4-7-low |
| medium | grok 4.7 (medium) | — |
| high | grok 4.7 (high) | grok-4-7-high |
| xhigh | grok 4.7 (xhigh) | grok-4.7-xhigh |