VECTOR WIREAI INTELLIGENCE
NVDA$1,847+3.2%MSFT$512+1.1%GOOGL$199-0.4%META$728+2.7%AMD$184-1.2%TSM$212+0.6%PLTR$98+4.1%AI IDX4,821+1.9%
PKT
SEEDRefresh Models Deals Regulatory Sources

DeepSeek-V4-Pro vs Kimi K2 (Reasoning)

DeepSeekvsMoonshot29 shared benchmarks263 head-to-head
BenchmarkDeepSeek-V4-ProKimi K2 (Reasoning)
AA Agentic Index63.347.9
AA Intelligence5333.5
AA-LCR7070.3
AA-Omniscience-10.6-21.4
Artificial Analysis Coding Index59.434.8
browsecomp83.460.2
critpt132.6
gdpval4924.5
GPQA Diamond90.584.5
HLE48.223.9
HLE (with tools)48.244.9
hmmt_nov_202594.489.2
IFBench76.568.1
imo_answer_bench89.878.6
LiveCodeBench93.579.2
MMLU-Pro87.584.6
OmniScience Accuracy4330.9
OmniScience Non-Hallucination12.225.8
scicode5044.8
simplebench50.939.6
SWE-bench Multilingual76.261.1
SWE-bench Verified80.671.3
Terminal-Bench 2.067.935.7
Terminal-Bench Hard46.231.1
vectara_answer_rate97.298.6
vectara_avg_summary_length153.859.2
vectara_factual_consistency91.482.1
vectara_hallucination_rate8.617.9
τ²-Bench Telecom (AA run)96.293

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.