VECTOR WIREAI INTELLIGENCE
PKT
Refresh Models Deals Regulatory Sources

Questions this page answers

Which is better, DeepSeek-V4-Pro or Qwen3.5 122B A10B?
Across 27 shared benchmarks, DeepSeek-V4-Pro scores higher on 24 and Qwen3.5 122B A10B on 3. The widest gap is AA Agentic Index, where DeepSeek-V4-Pro scores 63.3 against 21.3. Tracked API pricing per million tokens: DeepSeek-V4-Pro $0.66 in / $1.98 out, Qwen3.5 122B A10B $0.40 in / $3.20 out.

DeepSeek-V4-Pro vs Qwen3.5 122B A10B

Across 27 shared benchmarks, DeepSeek-V4-Pro scores higher on 24 and Qwen3.5 122B A10B on 3. The widest gap is AA Agentic Index, where DeepSeek-V4-Pro scores 63.3 against 21.3. Tracked API pricing per million tokens: DeepSeek-V4-Pro $0.66 in / $1.98 out, Qwen3.5 122B A10B $0.40 in / $3.20 out.

DeepSeekvsAlibaba27 shared benchmarks243 head-to-head
BenchmarkDeepSeek-V4-ProQwen3.5 122B A10B
AA Agentic Index63.321.3
AA Intelligence5332.8
AA-LCR7070.3
AA-Omniscience-10.6-41.5
Artificial Analysis Coding Index59.445.7
browsecomp83.463.8
coding_arena_elo15821358
critpt130.9
gdpval4924.3
GPQA Diamond90.586.6
HLE48.247.5
IFBench76.576.1
MMLU-Pro87.586.7
OmniScience Accuracy4324.4
OmniScience Non-Hallucination12.212.9
scicode5042
SWE-bench Verified80.672
TauBench V3 - Banking30.115.3
Terminal-Bench 2.067.949.4
Terminal-Bench 2.172.147.6
Terminal-Bench Hard46.231.1
vectara_answer_rate97.299.8
vectara_avg_summary_length153.886.4
vectara_factual_consistency91.488.8
vectara_hallucination_rate8.611.2
τ²-Bench Telecom (AA run)96.293.6
τ³-Bench25.813.6

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing. Quoted rates are the price-setter row we currently track for each model — its direct or vendor-official listing where one exists (deepseek-official, alibaba-official), otherwise the lowest tracked offer.