Gemma 4 26B A4B
Gemma 4 26B A4B is behind the leaders in instruction following, long context, multimodal tasks, and factuality. Too few results yet to rate reasoning, coding, agentic tasks, safety, math, or multilingual tasks.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Reasoning, Coding, Agentic, Safety, Math or Multilingual.
Price
$0.13input$0.40outputper million tokens
From Artificial Analysis · 5 providers tracked · All prices
Evidence
135results on99benchmarks
- 14 independently verified
- 34 aggregator
- 17 vendor-reported
- 70 cross-referenced
From 14 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingJSON modeReasoning
As listed by OpenRouter
Gemma 4 26B A4B benchmark results
135 results on 99 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
26.6% behind the leader1 of 3 ranked benchmarks measured
- IFBench72.45Oct 8, 2026aa_ifbench
31.0% behind the leader2 of 3 ranked benchmarks measured
- 65.67Oct 8, 2026
- MRCR v2 (8-needle, 128K)44.10Oct 8, 2026MRCR v2 8 needle 128k (average)
Show 1 more long context resultHide 1 long context result
- 42.33Oct 8, 2026
31.1% behind the leader4 of 6 ranked benchmarks measured
- 78.40Aug 24, 2026
- MathVista79.40Aug 24, 2026Mathvista(mini)
- MMMU-Pro69.25Oct 8, 2026aa_mmmu_pro
- CharXiv (reasoning)69.00Aug 24, 2026CharXiv(RQ)
Show 5 more multimodal resultsHide 5 multimodal results
- 1259.81Aug 25, 2026
- 1238Jun 17, 2026
- MMMU-Pro66.71Oct 8, 2026aa_mmmu_pro
- 73.80Oct 8, 2026
- 73.80Aug 24, 2026
34.4% behind the leader3 of 4 ranked benchmarks measured
- 5.20May 2, 2026
- AA-Omniscience · Accuracy19.10Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination13.56Oct 8, 2026omniscienceNonHallucination
Show 2 more factuality resultsHide 2 factuality results
- AA-Omniscience · Accuracy15.48Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination7.95Oct 8, 2026omniscienceNonHallucination
0 of 6 ranked benchmarks measured
- 0.00Oct 8, 2026
- GPQA Diamond79.19Oct 8, 2026gpqa
- GPQA Diamond71.41Oct 8, 2026gpqa
Show 8 more reasoning resultsHide 8 reasoning results
- 82.30Oct 8, 2026
- GPQA Diamond82.30Aug 24, 2026GPQA
- GPQA Diamond79.61Aug 11, 2026GPQA Diamond (no tools)
- Humanity's Last Exam11.49Oct 8, 2026aa_hle
- Humanity's Last Exam19.32Oct 8, 2026aa_hle
- Humanity's Last Exam8.70Oct 8, 2026HLE no tools
- 17.20Oct 7, 2026
- 8.70Jun 6, 2026
0 of 10 ranked benchmarks measured
- 1358.52May 22, 2026
- SciCode40.05Oct 8, 2026aa_scicode
- Terminal-Bench 2.138.95Oct 8, 2026terminalbenchV21
Show 12 more coding resultsHide 12 coding results
- 77.10Oct 8, 2026
- 77.10Jun 6, 2026
- SciCode37.27Sep 4, 2026aa_scicode
- 40.28Aug 11, 2026
- 43.40Aug 11, 2026
- 17.30May 3, 2026
- 13.80May 3, 2026
- 57.40Aug 11, 2026
- 17.40May 3, 2026
- 37.22Aug 11, 2026
- Terminal-Bench Hard25.00Oct 8, 2026aa_terminalbench_hard
- Terminal-Bench Hard13.64Oct 8, 2026aa_terminalbench_hard
0 of 7 ranked benchmarks measured
- 23.63Oct 8, 2026
- 3.44Oct 8, 2026
- τ-Bench V3 · Banking11.96Oct 8, 2026tauBanking
Show 3 more agentic resultsHide 3 agentic results
- 26.30Aug 11, 2026
- 25.72May 2, 2026
- 50.00Jun 6, 2026
0 of 5 ranked benchmarks measured
Show 1 more math resultHide 1 math result
- HMMT Feb 202679.00Jun 6, 2026HMMT Feb 26
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- 1438Jul 23, 2026
- AA Intelligence26.00Jul 3, 2026Artificial Analysis Intelligence Index
- AA Intelligence13.00Jun 20, 2026Artificial Analysis Intelligence Index
- 8.31Jun 9, 2026
- 12.90May 22, 2026
- 6.12May 22, 2026
Show 73 more resultsHide 73 results
- 11.04Sep 4, 2026
- 32.15May 2, 2026
- AA Intelligence13.13Oct 8, 2026aa_intelligence_index
- AA Intelligence16.67Oct 8, 2026aa_intelligence_index
- -62.32Oct 8, 2026
- -50.83Oct 8, 2026
- AGIEval-En55.28Aug 11, 2026AGIEval-EN (CoT)
- AI2D88.30Aug 24, 2026AI2D_TEST
- ARC-Challenge92.83Aug 11, 2026ARC-Challenge (25-shot)
- Artificial Analysis Coding Index39.32Sep 9, 2026aa_coding_index
- 22.44May 2, 2026
- 64.80Oct 8, 2026
- 82.50Jun 6, 2026
- 74.50Aug 24, 2026
- Claw Eval (pass@3)28.00Aug 24, 2026Claw-Eval Pass^3
- 58.80May 3, 2026
- 1718.00Oct 8, 2026
- 16.20Jun 6, 2026
- 6.29May 22, 2026
- 807.00Aug 11, 2026
- 74.78Aug 24, 2026
- 82.30Jun 6, 2026
- GSM8K77.03Aug 11, 2026GSM8K (8-shot, CoT)
- 66.10Aug 24, 2026
- 85.26Aug 11, 2026
- 17.42Aug 11, 2026
- HLE (with tools)17.20Oct 8, 2026HLE with search
- HMMT Feb. 202591.70Jun 6, 2026HMMT Feb 25
- HMMT Nov. 202587.50Jun 6, 2026HMMT Nov 25
- 50.00Aug 11, 2026
- 91.40Aug 9, 2026
- 82.40Oct 8, 2026
- MBPP68.39Aug 11, 2026MBPP (3-shot)
- 14.20Jun 6, 2026
- 58.10Oct 8, 2026
- Minerva Math43.74Aug 11, 2026Minerva Math (4-shot)
- 89.00Aug 24, 2026
- 77.81Aug 11, 2026
- 82.60Oct 8, 2026
- MMLU-Pro50.02Aug 11, 2026MMLU-Pro (5-shot)
- 85.20Aug 11, 2026
- 82.60Jun 6, 2026
- 92.70Jun 6, 2026
- 86.30Oct 8, 2026
- 11.60Jun 6, 2026
- 74.40Aug 24, 2026
- OmniDocBench 1.5 (average edit distance, lower is better)lower is better0.15Oct 8, 2026
- 48.80Aug 11, 2026
- 74.70Aug 11, 2026
- 83.84Aug 11, 2026
- PIQA (Acc.)82.21Aug 11, 2026PIQA (acc)
- 38.70Jun 6, 2026
- 1178.00Jun 6, 2026
- 72.20Aug 24, 2026
- 72.93Aug 11, 2026
- 85.73Aug 11, 2026
- 52.20Aug 24, 2026
- 12.30May 3, 2026
- 61.40Jun 6, 2026
- 85.50Oct 7, 2026
- 34.20May 3, 2026
- 12.00Jun 6, 2026
- vectara_answer_rate99.80May 2, 2026Answer Rate
- vectara_avg_summary_length67.10May 2, 2026Average Summary Length (Words)
- vectara_factual_consistency94.80May 2, 2026Factual Consistency Rate
- 36.90Jun 6, 2026
- 38.30Jun 6, 2026
- WinoGrande79.08Aug 11, 2026WinoGrande (5-shot)
- τ²-Bench68.20Oct 8, 2026Tau2 (average over 3)
- τ²-Bench Telecom (AA run)40.35Oct 8, 2026aa_tau2
- τ²-Bench Telecom (AA run)43.57Oct 8, 2026aa_tau2
- τ³-Bench59.00Jun 6, 2026TAU3-Bench
- τ³-Bench Banking14.02Aug 11, 2026τ³-bench (Banking)
Gemma 4 26B A4B: common questions
Who makes Gemma 4 26B A4B?
Gemma 4 26B A4B is made by Google.
When was Gemma 4 26B A4B released?
Gemma 4 26B A4B was released on Apr 2, 2026, according to Artificial Analysis.
What is Gemma 4 26B A4B good at?
Gemma 4 26B A4B is behind the leaders in instruction following, long context, multimodal tasks, and factuality. Too few results yet to rate reasoning, coding, agentic tasks, safety, math, or multilingual tasks.
How much does Gemma 4 26B A4B cost?
Gemma 4 26B A4B costs $0.13 per million input tokens and $0.40 per million output tokens, according to Artificial Analysis. We track its price at 5 providers. At a mix of three input tokens to one output token, it is cheaper than 82% of the 331 priced models we track.
How many benchmarks has Gemma 4 26B A4B been tested on?
We track 135 results for Gemma 4 26B A4B on 99 benchmarks from 14 sources, 14 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does Gemma 4 26B A4B support?
OpenRouter lists tool calling, json mode, and reasoning for Gemma 4 26B A4B.
About this record
Where Gemma 4 26B A4B's numbers come from, and every name it appears under.
- Tracked since
- May 2, 2026
- Newest source mention
- Oct 8, 2026
Where the results come from
Verification: 135 scores · 14 independently verified · 34 aggregator-attributed · 70 vendor cross-reference · 17 vendor-reported. How these tiers are assigned
From 14 sources on 6 sites. Hugging Face supplies 89 of them; the 14 independently verified results come from 5 sites. Bars are coloured by trust tier.
- huggingface.co89
- artificialanalysis.ai36
- raw.githubusercontent.com4
- api.llm-stats.com2
- datasets-server.huggingface.co2
- lmarena.ai2