Llama 3.1 Nemotron Instruct 70B
Llama 3.1 Nemotron Instruct 70B is behind the leaders in factuality. Too few results yet to rate reasoning, coding, agentic tasks, safety, long context, math, multimodal tasks, multilingual tasks, or instruction following.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Reasoning, Coding, Agentic, Safety, Long Context, Math, Multimodal, Multilingual or Instruction Following.
Price
No current price is tracked for this model. See the rate card
Evidence
27results on26benchmarks
- 3 independently verified
- 15 aggregator
- 9 vendor-reported
From 4 sources · latest Oct 7, 2026 · How verification works
API features
Tool callingJSON mode
As listed by OpenRouter
Llama 3.1 Nemotron Instruct 70B benchmark results
27 results on 26 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
36.7% behind the leader2 of 4 ranked benchmarks measured
- AA-Omniscience · Non-hallucination28.74Oct 7, 2026omniscienceNonHallucination
- AA-Omniscience · Accuracy17.77Oct 7, 2026omniscienceAccuracy
0 of 6 ranked benchmarks measured
- 0.00Oct 7, 2026
- GPQA Diamond46.46Oct 7, 2026gpqa
- Humanity's Last Exam4.19Oct 7, 2026aa_hle
0 of 10 ranked benchmarks measured
- LiveBench · Coding32.81Aug 23, 2026livebench_coding@2025-04-07
- Terminal-Bench Hard4.55Oct 7, 2026aa_terminalbench_hard
- SciCode23.26Sep 4, 2026aa_scicode
0 of 7 ranked benchmarks measured
- 0.00Jun 15, 2026
0 of 3 ranked benchmarks measured
- 8.33Oct 7, 2026
0 of 3 ranked benchmarks measured
- LiveBench · Instruction Following68.40Aug 23, 2026livebench_instruction_following@2025-04-07
- IFBench30.75Oct 7, 2026aa_ifbench
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- livebench_language34.17Aug 23, 2026livebench_language@2025-04-07
- τ²-Bench Telecom (AA run)23.10Oct 7, 2026aa_tau2
- AA Intelligence6.94Oct 7, 2026aa_intelligence_index
- -40.83Oct 7, 2026
- 10.78Jun 18, 2026
- 7.70Jun 18, 2026
Show 9 more resultsHide 9 results
- AlpacaEval 2 LC57.60Sep 10, 2026
- ARC-Challenge69.20Oct 6, 2026ARC-C
- 85.00Sep 10, 2026
- 91.43Oct 6, 2026
- 85.58Oct 6, 2026
- 0.09Oct 6, 2026
- MT-Bench8.98Sep 10, 2026MT-Bench (GPT-4-Turbo)
- 58.63Oct 6, 2026
- 84.53Oct 6, 2026
Llama 3.1 Nemotron Instruct 70B: common questions
Who makes Llama 3.1 Nemotron Instruct 70B?
Llama 3.1 Nemotron Instruct 70B is made by NVIDIA.
When was Llama 3.1 Nemotron Instruct 70B released?
Llama 3.1 Nemotron Instruct 70B was released on Oct 15, 2024, according to Artificial Analysis.
What is Llama 3.1 Nemotron Instruct 70B good at?
Llama 3.1 Nemotron Instruct 70B is behind the leaders in factuality. Too few results yet to rate reasoning, coding, agentic tasks, safety, long context, math, multimodal tasks, multilingual tasks, or instruction following.
How many benchmarks has Llama 3.1 Nemotron Instruct 70B been tested on?
We track 27 results for Llama 3.1 Nemotron Instruct 70B on 26 benchmarks from 4 sources, 3 of them independently verified. The latest was recorded on Oct 7, 2026.
Which API features does Llama 3.1 Nemotron Instruct 70B support?
OpenRouter lists tool calling and json mode for Llama 3.1 Nemotron Instruct 70B.
About this record
Where Llama 3.1 Nemotron Instruct 70B's numbers come from, and every name it appears under.
- Tracked since
- Apr 25, 2026
- Newest source mention
- May 18, 2026
Where the results come from
Verification: 27 scores · 3 independently verified · 15 aggregator-attributed · 9 vendor-reported. How these tiers are assigned
From 4 sources on 3 sites. Artificial Analysis supplies 15 of them; the 3 independently verified results come from 1 site. Bars are coloured by trust tier.
- artificialanalysis.ai15
- api.llm-stats.com6
- huggingface.co6