VECTOR WIREAI INTELLIGENCE
NVDA$1,847+3.2%MSFT$512+1.1%GOOGL$199-0.4%META$728+2.7%AMD$184-1.2%TSM$212+0.6%PLTR$98+4.1%AI IDX4,821+1.9%
PKT
SEEDRefresh Models Deals Regulatory Sources

Gemini 3.1 Pro vs Qwen3.7 Max

GooglevsAlibaba50 shared benchmarks3119 head-to-head
BenchmarkGemini 3.1 ProQwen3.7 Max
AA Agentic Index2330.9
AA Intelligence47.746.7
AA-LCR7974.7
AA-Omniscience32.913.5
aime_202698.397
arena_elo14861474
arena_text_factuality14711479
Artificial Analysis Coding Index68.866
coding_arena_elo14471517
critpt17.713.4
DeepSWE1018
DeepSWE 1.11221.6
Finance Agent v24348.4
gdpval23.338.5
GPQA Diamond94.392.4
HLE51.441.4
HLE (with tools)51.653.5
hmmt_feb_202694.797.1
hmmt_nov_202594.895
IFBench77.180.5
imo_answer_bench9190
itbenchSre30.342.5
livebench79.974.3
livebench_agentic_coding44.143.6
livebench_coding76.574.2
livebench_data_analysis78.571.8
livebench_instruction_following79.174
livebench_language85.479.7
livebench_math9185.3
livebench_reasoning8483.3
LiveCodeBench v691.791.6
MCP Atlas78.276.4
mcpmark55.960.8
MMLU-Pro9189.6
mmmlu92.690.3
nl2repo33.447.2
OmniScience Accuracy55.331.1
OmniScience Non-Hallucination50.174.4
scicode5953.5
simplebench79.670.4
simpleqa_verified77.358.5
SWE-bench Multilingual76.978.3
SWE-bench Pro54.260.6
SWE-bench Verified80.680.4
TauBench V3 - Banking21.411.8
Terminal-Bench 2.068.569.7
Terminal-Bench 2.17475
Terminal-Bench Hard68.550.8
τ²-Bench Telecom (AA run)99.394.7
τ³-Bench67.110.9

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.