Gemma 3 27B
Gemma 3 27B is behind the leaders in multimodal tasks. Too few results yet to rate reasoning, coding, agentic tasks, safety, long context, math, multilingual tasks, instruction following, or factuality.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Reasoning, Coding, Agentic, Safety, Long Context, Math, Multilingual, Instruction Following or Factuality.
Price
$0.08input$0.16outputper million tokens
From deepinfra · 2 providers tracked · All prices
Evidence
82results on66benchmarks
- 13 independently verified
- 19 aggregator
- 28 vendor-reported
- 22 cross-referenced
From 15 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingStructured outputs
As listed by OpenRouter
Research
17 papers reference Gemma 3 27BGemma 3 27B benchmark results
82 results on 66 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
47.2% behind the leader3 of 6 ranked benchmarks measured
- MMMU64.90Jun 5, 2026MMMU (val, Pass@1)
- MathVista67.60Oct 7, 2026MathVista-Mini
- MMMU-Pro48.03Oct 8, 2026aa_mmmu_pro
Show 5 more multimodal resultsHide 5 multimodal results
0 of 6 ranked benchmarks measured
- 0.00Oct 8, 2026
- GPQA Diamond42.83Oct 8, 2026gpqa
- Humanity's Last Exam4.40Oct 8, 2026aa_hle
Show 3 more reasoning resultsHide 3 reasoning results
- 42.40Oct 8, 2026
- 42.40Aug 25, 2026
- 46.00Jun 5, 2026
0 of 10 ranked benchmarks measured
- 11.38Oct 8, 2026
- LiveBench · Coding39.84Aug 23, 2026livebench_coding@2025-04-07
- SciCode23.26Oct 8, 2026aa_scicode
Show 3 more coding resultsHide 3 coding results
- 29.10Oct 8, 2026
- Terminal-Bench 2.14.49Oct 8, 2026terminalbenchV21
- Terminal-Bench Hard3.79Oct 8, 2026aa_terminalbench_hard
0 of 7 ranked benchmarks measured
- 0.00Oct 8, 2026
- Terminal-Bench 4.00.00Oct 8, 2026
- τ-Bench V3 · Banking0.82Oct 8, 2026tauBanking
0 of 3 ranked benchmarks measured
- 7.33Oct 8, 2026
- MRCR v2 (8-needle, 128K)13.50Oct 8, 2026MRCR v2 8 needle 128k (average)
0 of 5 ranked benchmarks measured
- AIME 202620.80Oct 8, 2026AIME 2026 no tools
0 of 3 ranked benchmarks measured
- LiveBench · Instruction Following71.50Aug 23, 2026livebench_instruction_following@2025-04-07
- IFBench31.84Oct 8, 2026aa_ifbench
0 of 4 ranked benchmarks measured
- 7.40May 2, 2026
- AA-Omniscience · Accuracy12.95Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination7.91Oct 8, 2026omniscienceNonHallucination
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- livebench_language40.39Aug 23, 2026livebench_language@2025-04-07
- AA Intelligence7.00Jul 10, 2026Artificial Analysis Intelligence Index
- vectara_avg_summary_length96.40May 2, 2026Average Summary Length (Words)
- vectara_answer_rate98.80May 2, 2026Answer Rate
- vectara_factual_consistency92.60May 2, 2026Factual Consistency Rate
- 4.90May 1, 2026
Show 45 more resultsHide 45 results
- 0.14Sep 9, 2026
- AA Intelligence4.85Oct 8, 2026aa_intelligence_index
- -67.22Oct 8, 2026
- 84.50Oct 7, 2026
- Artificial Analysis Coding Index10.06Sep 9, 2026aa_coding_index
- BBH87.60Oct 7, 2026BIG-Bench Hard
- 19.30Oct 8, 2026
- 78.00Oct 7, 2026
- 110.00Oct 8, 2026
- 86.60Oct 7, 2026
- 75.10Oct 7, 2026
- 95.90Oct 7, 2026
- 87.80Oct 7, 2026
- 87.80Aug 25, 2026
- 90.40Aug 31, 2026
- 90.40Jun 15, 2026
- 70.60Oct 7, 2026
- 29.70Aug 23, 2026
- 89.00Aug 25, 2026
- 46.00Oct 8, 2026
- MATH-Vision35.40Jun 5, 2026MATH-Vision (Pass@1)
- 35.50Jun 5, 2026
- 74.40Oct 7, 2026
- 74.40Aug 25, 2026
- 78.90Jun 5, 2026
- 76.90Aug 25, 2026
- 67.60Oct 8, 2026
- 67.50Oct 7, 2026
- 67.50Aug 25, 2026
- 56.60May 26, 2025
- 70.70Oct 8, 2026
- MMMU (val) (Pass@1)64.90Oct 7, 2026MMMU (val)
- 64.80Jun 5, 2026
- 63.10Jun 5, 2026
- 71.00Jun 5, 2026
- MMVU61.30Jun 5, 2026MMVU (Pass@1)
- 753.00Jun 5, 2026
- OmniDocBench 1.5 (average edit distance, lower is better)lower is better0.36Oct 8, 2026
- 62.50Jun 5, 2026
- 10.00Oct 7, 2026
- 10.00Jun 15, 2026
- 65.10Oct 7, 2026
- VideoMMMU61.80Jun 5, 2026VideoMMMU (Pass@1)
- τ²-Bench16.20Oct 8, 2026Tau2 (average over 3)
- τ²-Bench Telecom (AA run)10.53Oct 8, 2026aa_tau2
Gemma 3 27B: common questions
Who makes Gemma 3 27B?
Gemma 3 27B is made by Google.
When was Gemma 3 27B released?
Gemma 3 27B was released on Mar 12, 2025, according to Artificial Analysis.
What is Gemma 3 27B good at?
Gemma 3 27B is behind the leaders in multimodal tasks. Too few results yet to rate reasoning, coding, agentic tasks, safety, long context, math, multilingual tasks, instruction following, or factuality.
How much does Gemma 3 27B cost?
Gemma 3 27B costs $0.08 per million input tokens and $0.16 per million output tokens, according to deepinfra. We track its price at 2 providers. At a mix of three input tokens to one output token, it is cheaper than 91% of the 331 priced models we track.
How many benchmarks has Gemma 3 27B been tested on?
We track 82 results for Gemma 3 27B on 66 benchmarks from 15 sources, 13 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does Gemma 3 27B support?
OpenRouter lists tool calling and structured outputs for Gemma 3 27B.
About this record
Where Gemma 3 27B's numbers come from, and every name it appears under.
- Tracked since
- Jun 20, 2026
- Newest source mention
- Sep 28, 2026
Where the results come from
Verification: 82 scores · 13 independently verified · 19 aggregator-attributed · 22 vendor cross-reference · 28 vendor-reported. How these tiers are assigned
From 15 sources on 9 sites. Hugging Face supplies 37 of them; the 13 independently verified results come from 8 sites. Bars are coloured by trust tier.
- huggingface.co37
- artificialanalysis.ai20
- api.llm-stats.com16
- raw.githubusercontent.com4
- aider.chat1
- arxiv.org1
- datasets-server.huggingface.co1
- labs.scale.com1
- lmarena.ai1