Grok 4.20 (Beta Reasoning)
Grok 4.20 (Beta Reasoning) has too few ranked results yet to rate it on any capability.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability.
No capability has enough ranked results to rate yet.
Price
No current price is tracked for this model. See the rate card
Evidence
6results on5benchmarks
- 5 independently verified
- 1 vendor-reported
From 5 sources · latest Oct 7, 2026 · How verification works
Grok 4.20 (Beta Reasoning) benchmark results
6 results on 5 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
0 of 6 ranked benchmarks measured
- 0.09Sep 20, 2026
0 of 10 ranked benchmarks measured
- 1374.60Sep 20, 2026
0 of 6 ranked benchmarks measured
- 1262.59Sep 20, 2026
- 1251May 25, 2026
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- 1472Oct 5, 2026
- 28.49Oct 7, 2026
Grok 4.20 (Beta Reasoning): common questions
Who makes Grok 4.20 (Beta Reasoning)?
Grok 4.20 (Beta Reasoning) is made by SpaceXAI.
How many benchmarks has Grok 4.20 (Beta Reasoning) been tested on?
We track 6 results for Grok 4.20 (Beta Reasoning) on 5 benchmarks from 5 sources, 5 of them independently verified. The latest was recorded on Oct 7, 2026.
About this record
Where Grok 4.20 (Beta Reasoning)'s numbers come from, and every name it appears under.
- Tracked since
- May 10, 2026
- Newest source mention
- May 10, 2026
Where the results come from
Verification: 6 scores · 5 independently verified · 1 vendor-reported. How these tiers are assigned
From 5 sources on 4 sites. datasets-server.huggingface.co supplies 2 of them; the 5 independently verified results come from 3 sites. Bars are coloured by trust tier.
- datasets-server.huggingface.co2
- lmarena.ai2
- api.llm-stats.com1
- arcprize.org1