DeepSeek-V3.2-Speciale
DeepSeek-V3.2-Speciale is behind the leaders in long context, reasoning, coding, instruction following, and factuality. Too few results yet to rate agentic tasks, safety, math, multimodal tasks, or multilingual tasks.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Agentic, Safety, Math, Multimodal or Multilingual.
Price
$0.29input$0.43outputper million tokens
From deepseek · All prices
Evidence
27results on25benchmarks
- 5 independently verified
- 15 aggregator
- 7 vendor-reported
From 7 sources · latest Oct 8, 2026 · How verification works
API features
Structured outputsReasoning
As listed by OpenRouter
DeepSeek-V3.2-Speciale benchmark results
27 results on 25 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
25.3% behind the leader1 of 3 ranked benchmarks measured
- 70.00Oct 8, 2026
25.9% behind the leader4 of 6 ranked benchmarks measured
- GPQA Diamond87.07Oct 8, 2026gpqa
- 52.60May 10, 2026
- Humanity's Last Exam28.73Oct 8, 2026aa_hle
- 7.43Oct 8, 2026
Show 1 more reasoning resultHide 1 reasoning result
- 30.60Oct 7, 2026
29.5% behind the leader3 of 10 ranked benchmarks measured
- 73.10Oct 7, 2026
- SciCode43.98Sep 4, 2026aa_scicode
- Terminal-Bench Hard34.85Oct 8, 2026aa_terminalbench_hard
32.3% behind the leader1 of 3 ranked benchmarks measured
- IFBench63.88Oct 8, 2026aa_ifbench
34.4% behind the leader2 of 4 ranked benchmarks measured
- AA-Omniscience · Accuracy37.77Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination10.85Oct 8, 2026omniscienceNonHallucination
0 of 7 ranked benchmarks measured
- 0.00Jun 15, 2026
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- 59.17Sep 2, 2026
- 97.50May 11, 2026
- 93.33May 10, 2026
- 95.83May 2, 2026
- AA Intelligence14.46Oct 8, 2026aa_intelligence_index
- -17.72Oct 8, 2026
Show 8 more resultsHide 8 results
- 0.00Jun 18, 2026
- 96.00Oct 7, 2026
- 37.89Jun 18, 2026
- 99.20Oct 7, 2026
- 80.30Oct 7, 2026
- 46.40Oct 7, 2026
- 35.20Oct 7, 2026
- τ²-Bench Telecom (AA run)0.00Oct 8, 2026aa_tau2
DeepSeek-V3.2-Speciale: common questions
Who makes DeepSeek-V3.2-Speciale?
DeepSeek-V3.2-Speciale is made by DeepSeek.
When was DeepSeek-V3.2-Speciale released?
DeepSeek-V3.2-Speciale was released on Dec 1, 2025, according to Artificial Analysis.
What is DeepSeek-V3.2-Speciale good at?
DeepSeek-V3.2-Speciale is behind the leaders in long context, reasoning, coding, instruction following, and factuality. Too few results yet to rate agentic tasks, safety, math, multimodal tasks, or multilingual tasks.
How much does DeepSeek-V3.2-Speciale cost?
DeepSeek-V3.2-Speciale costs $0.29 per million input tokens and $0.43 per million output tokens, according to deepseek. At a mix of three input tokens to one output token, it is cheaper than 70% of the 331 priced models we track.
How many benchmarks has DeepSeek-V3.2-Speciale been tested on?
We track 27 results for DeepSeek-V3.2-Speciale on 25 benchmarks from 7 sources, 5 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does DeepSeek-V3.2-Speciale support?
OpenRouter lists structured outputs and reasoning for DeepSeek-V3.2-Speciale.
About this record
Where DeepSeek-V3.2-Speciale's numbers come from, and every name it appears under.
- Tracked since
- Apr 27, 2026
- Newest source mention
- Aug 23, 2026
Where the results come from
Verification: 27 scores · 5 independently verified · 15 aggregator-attributed · 7 vendor-reported. How these tiers are assigned
From 7 sources on 4 sites. Artificial Analysis supplies 15 of them; the 5 independently verified results come from 2 sites. Bars are coloured by trust tier.
- artificialanalysis.ai15
- api.llm-stats.com7
- matharena.ai4
- simple-bench.com1