Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

Llama 4 Scout

Score basis

Llama 4 Scout is behind the leaders in multimodal tasks. Too few results yet to rate reasoning, coding, agentic tasks, safety, long context, math, multilingual tasks, instruction following, or factuality.

Capability profile

Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.

Limited1 capability
  1. Multimodal−42.2%3 of 6
Not rated9 capabilities

Too few results yet to rate Reasoning, Coding, Agentic, Safety, Long Context, Math, Multilingual, Instruction Following or Factuality.

Price

$0.19input$0.68outputper million tokens

From Artificial Analysis · 5 providers tracked · All prices

Cheaper than 71% of 331 priced models · 3:1 input-to-output blend, log scale

Evidence

46results on43benchmarks

  • 15 independently verified
  • 19 aggregator
  • 12 vendor-reported

From 14 sources · latest Oct 8, 2026 · How verification works

API features

Tool callingStructured outputs

As listed by OpenRouter

Llama 4 Scout benchmark results

46 results on 43 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.

Multimodal

LimitedFull multimodal ranking

42.2% behind the leader3 of 6 ranked benchmarks measured

Show 1 more multimodal resultHide 1 multimodal result

Reasoning

Not enough dataFull reasoning ranking

0 of 6 ranked benchmarks measured

Show 2 more reasoning resultsHide 2 reasoning results

Coding

Not enough dataFull coding ranking

0 of 10 ranked benchmarks measured

Show 1 more coding resultHide 1 coding result

Agentic

Not enough dataFull agentic ranking

0 of 7 ranked benchmarks measured

Long Context

Not enough dataFull long context ranking

0 of 3 ranked benchmarks measured

Instruction Following

Not enough dataFull instruction following ranking

0 of 3 ranked benchmarks measured

Factuality

Not enough dataFull factuality ranking

0 of 4 ranked benchmarks measured

More results

Benchmarks outside the capability baskets. They are not ranked against a leader.

Show 19 more resultsHide 19 results

Llama 4 Scout: common questions

Who makes Llama 4 Scout?

Llama 4 Scout is made by Meta.

When was Llama 4 Scout released?

Llama 4 Scout was released on Apr 5, 2025, according to Artificial Analysis.

What is Llama 4 Scout good at?

Llama 4 Scout is behind the leaders in multimodal tasks. Too few results yet to rate reasoning, coding, agentic tasks, safety, long context, math, multilingual tasks, instruction following, or factuality.

How much does Llama 4 Scout cost?

Llama 4 Scout costs $0.19 per million input tokens and $0.68 per million output tokens, according to Artificial Analysis. We track its price at 5 providers. At a mix of three input tokens to one output token, it is cheaper than 71% of the 331 priced models we track.

How many benchmarks has Llama 4 Scout been tested on?

We track 46 results for Llama 4 Scout on 43 benchmarks from 14 sources, 15 of them independently verified. The latest was recorded on Oct 8, 2026.

Which API features does Llama 4 Scout support?

OpenRouter lists tool calling and structured outputs for Llama 4 Scout.

About this record

Where Llama 4 Scout's numbers come from, and every name it appears under.

Tracked since
May 1, 2026
Newest source mention
Aug 31, 2026

Where the results come from

Verification: 46 scores · 15 independently verified · 19 aggregator-attributed · 12 vendor-reported. How these tiers are assigned

From 14 sources on 6 sites. Artificial Analysis supplies 20 of them; the 15 independently verified results come from 5 sites. Bars are coloured by trust tier.

  • artificialanalysis.ai20
  • api.llm-stats.com9
  • raw.githubusercontent.com7
  • storage.googleapis.com6
  • arcprize.org2
  • swebench.com2

Also known as

Llama 4 Scout (17Bx16E) InstructLlama 4 Scout (Non-Reasoning)Llama 4 Scout (Reasoning)Llama 4 Scout 17B 16E Instructllama 4 scout instructllama4-scout