VECTOR WIREAI INTELLIGENCE
NVDA$1,847+3.2%MSFT$512+1.1%GOOGL$199-0.4%META$728+2.7%AMD$184-1.2%TSM$212+0.6%PLTR$98+4.1%AI IDX4,821+1.9%
PKT
SEEDRefresh Models Deals Regulatory Sources

Gemini 3 Pro vs Qwen3.5 397B A17B

GooglevsAlibaba46 shared benchmarks2917 head-to-head
BenchmarkGemini 3 ProQwen3.5 397B A17B
AA Agentic Index5253.3
AA Intelligence40.634.3
AA-LCR7372.7
AA-Omniscience15.3-30.8
aime_202691.794.2
apexAgents18.415.3
arena_elo14851442
arena_text_factuality14811445
arena_vision12901249
Artificial Analysis Coding Index46.548.2
browsecomp59.269
BrowseComp (context management)67.678.6
browsecomp_zh66.870.3
charxiv_rq81.480.8
coding_arena_elo14381399
critpt9.11.7
gdpval34.235.8
GPQA Diamond91.989.3
HLE45.829
HLE (with tools)45.848.3
hmmt_feb_202597.594.8
hmmt_feb_202686.487.9
hmmt_nov_202593.392.7
IFBench70.478.8
imo_answer_bench83.380.9
LiveCodeBench v690.783.6
longbench_v268.263.2
MLVU80.786.7
MMLU-Pro90.187.8
mmmlu91.888.5
MMMU-Pro8179
multichallenge65.767.6
OmniScience Accuracy55.830.8
OmniScience Non-Hallucination1017.3
scicode56.142
Seal-045.546.9
simpleqa_verified72.926
SimpleVQA69.767.1
SWE-bench Multilingual68.769.3
SWE-bench Pro43.350.9
SWE-bench Verified7876.4
Terminal-Bench 2.054.252.5
Terminal-Bench Hard56.940.9
toolathlon36.438.3
VideoMMMU87.684.7
τ²-Bench Telecom (AA run)9895.6

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.