Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

Llama 3.1 Instruct 405B

Score basis

Llama 3.1 Instruct 405B is behind the leaders in factuality. Too few results yet to rate reasoning, coding, agentic tasks, safety, long context, math, multimodal tasks, multilingual tasks, or instruction following.

Capability profile

Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.

Limited1 capability
  1. Factuality−27.8%2 of 4
Not rated9 capabilities

Too few results yet to rate Reasoning, Coding, Agentic, Safety, Long Context, Math, Multimodal, Multilingual or Instruction Following.

Price

No current price is tracked for this model. See the rate card

Evidence

151results on110benchmarks

  • 17 independently verified
  • 15 aggregator
  • 21 vendor-reported
  • 98 cross-referenced

From 24 sources · latest Oct 7, 2026 · How verification works

Llama 3.1 Instruct 405B benchmark results

151 results on 110 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.

Factuality

LimitedFull factuality ranking

27.8% behind the leader2 of 4 ranked benchmarks measured

Reasoning

Not enough dataFull reasoning ranking

0 of 6 ranked benchmarks measured

Show 5 more reasoning resultsHide 5 reasoning results

Coding

Not enough dataFull coding ranking

0 of 10 ranked benchmarks measured

Show 2 more coding resultsHide 2 coding results

Agentic

Not enough dataFull agentic ranking

0 of 7 ranked benchmarks measured

Long Context

Not enough dataFull long context ranking

0 of 3 ranked benchmarks measured

Instruction Following

Not enough dataFull instruction following ranking

0 of 3 ranked benchmarks measured

More results

Benchmarks outside the capability baskets. They are not ranked against a leader.

Show 125 more resultsHide 125 results

Llama 3.1 Instruct 405B: common questions

Who makes Llama 3.1 Instruct 405B?

Llama 3.1 Instruct 405B is made by Meta.

When was Llama 3.1 Instruct 405B released?

Llama 3.1 Instruct 405B was released on Jul 23, 2024, according to Artificial Analysis.

What is Llama 3.1 Instruct 405B good at?

Llama 3.1 Instruct 405B is behind the leaders in factuality. Too few results yet to rate reasoning, coding, agentic tasks, safety, long context, math, multimodal tasks, multilingual tasks, or instruction following.

How many benchmarks has Llama 3.1 Instruct 405B been tested on?

We track 151 results for Llama 3.1 Instruct 405B on 110 benchmarks from 24 sources, 17 of them independently verified. The latest was recorded on Oct 7, 2026.

About this record

Where Llama 3.1 Instruct 405B's numbers come from, and every name it appears under.

Tracked since
Apr 27, 2026
Newest source mention
Sep 10, 2026

Where the results come from

Verification: 151 scores · 17 independently verified · 15 aggregator-attributed · 98 vendor cross-reference · 21 vendor-reported. How these tiers are assigned

From 24 sources on 8 sites. raw.githubusercontent.com supplies 59 of them; the 17 independently verified results come from 4 sites. Bars are coloured by trust tier.

  • raw.githubusercontent.com59
  • huggingface.co36
  • arxiv.org18
  • artificialanalysis.ai15
  • storage.googleapis.com12
  • api.llm-stats.com9
  • labs.scale.com1
  • simple-bench.com1

Also known as

llama 3.1 405bllama 3.1 405b instructllama 3.1 instruct turbo (405b)llama-3-1-instruct-405bllama-3-1-instruct-405b (Non-Reasoning)llama-3-1-instruct-405b (reasoning)llama-3.1 405b-inst.llama3-1-405b-instructllama3.1 405bllama3.1 405b inst.meta-llama-3.1-405b-instruct-turbo