Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

DeepSeek-V3.1-Terminus

Score basis

DeepSeek-V3.1-Terminus is behind the leaders in long context, factuality, coding, reasoning, instruction following, and agentic tasks. Too few results yet to rate safety, math, multimodal tasks, or multilingual tasks.

Capability profile

Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.

Limited6 capabilities
  1. Long Context−26.2%1 of 3
  2. Factuality−29.0%2 of 4
  3. Coding−35.1%3 of 10
  4. Reasoning−35.2%3 of 6
  5. Instruction Following−36.0%2 of 3
  6. Agentic−42.3%3 of 7
Not rated4 capabilities

Too few results yet to rate Safety, Math, Multimodal or Multilingual.

Price

$0.27input$1.00outputper million tokens

From Artificial Analysis · 2 providers tracked · All prices

Cheaper than 60% of 330 priced models · 3:1 input-to-output blend, log scale

Evidence

37results on21benchmarks

  • 1 independently verified
  • 32 aggregator
  • 4 cross-referenced

From 3 sources · latest Oct 8, 2026 · How verification works

API features

Tool callingStructured outputsReasoning

As listed by OpenRouter

DeepSeek-V3.1-Terminus benchmark results

37 results on 21 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.

Long Context

LimitedFull long context ranking

26.2% behind the leader1 of 3 ranked benchmarks measured

Show 1 more long context resultHide 1 long context result

Factuality

LimitedFull factuality ranking

29.0% behind the leader2 of 4 ranked benchmarks measured

Show 2 more factuality resultsHide 2 factuality results

Coding

LimitedFull coding ranking

35.1% behind the leader3 of 10 ranked benchmarks measured

Show 2 more coding resultsHide 2 coding results

Reasoning

LimitedFull reasoning ranking

35.2% behind the leader3 of 6 ranked benchmarks measured

Show 4 more reasoning resultsHide 4 reasoning results

Instruction Following

LimitedFull instruction following ranking

36.0% behind the leader2 of 3 ranked benchmarks measured

Show 1 more instruction following resultHide 1 instruction following result

Agentic

LimitedFull agentic ranking

42.3% behind the leader3 of 7 ranked benchmarks measured

Show 1 more agentic resultHide 1 agentic result

More results

Benchmarks outside the capability baskets. They are not ranked against a leader.

Show 6 more resultsHide 6 results

DeepSeek-V3.1-Terminus: common questions

Who makes DeepSeek-V3.1-Terminus?

DeepSeek-V3.1-Terminus is made by DeepSeek.

When was DeepSeek-V3.1-Terminus released?

DeepSeek-V3.1-Terminus was released on Sep 22, 2025, according to Artificial Analysis.

What is DeepSeek-V3.1-Terminus good at?

DeepSeek-V3.1-Terminus is behind the leaders in long context, factuality, coding, reasoning, instruction following, and agentic tasks. Too few results yet to rate safety, math, multimodal tasks, or multilingual tasks.

How much does DeepSeek-V3.1-Terminus cost?

DeepSeek-V3.1-Terminus costs $0.27 per million input tokens and $1.00 per million output tokens, according to Artificial Analysis. We track its price at 2 providers. At a mix of three input tokens to one output token, it is cheaper than 60% of the 330 priced models we track.

How many benchmarks has DeepSeek-V3.1-Terminus been tested on?

We track 37 results for DeepSeek-V3.1-Terminus on 21 benchmarks from 3 sources, 1 of them independently verified. The latest was recorded on Oct 8, 2026.

Which API features does DeepSeek-V3.1-Terminus support?

OpenRouter lists tool calling, structured outputs, and reasoning for DeepSeek-V3.1-Terminus.

About this record

Where DeepSeek-V3.1-Terminus's numbers come from, and every name it appears under.

Tracked since
Jun 18, 2026
Newest source mention
Sep 10, 2026

Where the results come from

Verification: 37 scores · 1 independently verified · 32 aggregator-attributed · 4 vendor cross-reference. How these tiers are assigned

From 3 sources on 2 sites. Artificial Analysis supplies 33 of them; the 1 independently verified result comes from 1 site. Bars are coloured by trust tier.

  • artificialanalysis.ai33
  • huggingface.co4

Also known as

deepseek v3.1 terminus (non-reasoning)deepseek v3.1 terminus (reasoning)deepseek-v3-1-terminusdeepseek-v3-1-terminus-reasoning