VECTOR WIREAI INTELLIGENCE
NVDA$1,847+3.2%MSFT$512+1.1%GOOGL$199-0.4%META$728+2.7%AMD$184-1.2%TSM$212+0.6%PLTR$98+4.1%AI IDX4,821+1.9%
PKT
SEEDRefresh Models Deals Regulatory Sources

DeepSeek-V4-Pro vs Qwen3.7 Max

DeepSeekvsAlibaba45 shared benchmarks1233 head-to-head
BenchmarkDeepSeek-V4-ProQwen3.7 Max
AA Agentic Index63.330.9
AA Intelligence5346.7
AA-LCR7074.7
AA-Omniscience-10.613.5
aime_202696.797
arena_elo14571474
arena_text_factuality14501479
Artificial Analysis Coding Index59.466
coding_arena_elo15821517
critpt1313.4
DeepSWE12.818
gdpval4938.5
GPQA Diamond90.592.4
HLE48.241.4
HLE (with tools)48.253.5
hmmt_feb_202695.297.1
hmmt_nov_202594.495
IFBench76.580.5
imo_answer_bench89.890
itbenchSre38.342.5
livebench73.674.3
livebench_agentic_coding42.643.6
livebench_coding7074.2
livebench_data_analysis74.571.8
livebench_instruction_following62.474
livebench_language78.179.7
livebench_math90.785.3
livebench_reasoning82.783.3
MCP Atlas74.276.4
MMLU-Pro87.589.6
nl2repo38.547.2
OmniScience Accuracy4331.1
OmniScience Non-Hallucination12.274.4
scicode5053.5
simplebench50.970.4
simpleqa_verified57.958.5
SWE-bench Multilingual76.278.3
SWE-bench Pro55.460.6
SWE-bench Verified80.680.4
TauBench V3 - Banking30.111.8
Terminal-Bench 2.067.969.7
Terminal-Bench 2.172.175
Terminal-Bench Hard46.250.8
τ²-Bench Telecom (AA run)96.294.7
τ³-Bench25.810.9

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.