VECTOR WIREAI INTELLIGENCE
NVDA$1,847+3.2%MSFT$512+1.1%GOOGL$199-0.4%META$728+2.7%AMD$184-1.2%TSM$212+0.6%PLTR$98+4.1%AI IDX4,821+1.9%
PKT
SEEDRefresh Models Deals Regulatory Sources

Qwen3.5 397B A17B vs Qwen3.6 35B A3B

AlibabavsAlibaba55 shared benchmarks478 head-to-head
BenchmarkQwen3.5 397B A17BQwen3.6 35B A3B
AA Agentic Index53.321.6
AA Intelligence34.332.1
AA-LCR72.766.7
AA-Omniscience-30.8-22.2
aime_202694.292.7
Artificial Analysis Coding Index48.241.9
browsecomp6948.7
C-Eval9390
CC-OCR8281.9
charxiv_rq80.878
Claw Eval (pass@3)48.150
Claw-Eval Avg70.768.7
critpt1.70.3
DeepPlanning34.325.9
gdpval35.827.8
GDPval-AA v29621015
GPQA Diamond89.386
HLE2922.2
HMMT 202592.789.1
hmmt_feb_202594.890.7
hmmt_feb_202687.983.6
hmmt_nov_202592.789.1
IFBench78.864.4
imo_answer_bench80.978.9
LiveCodeBench v683.680.4
mcpmark46.137
MLVU86.786.2
mmlu_redux94.993.3
MMLU-Pro87.885.6
MMMU8581.7
MMMU-Pro7975.3
MV-Bench77.674.6
nl2repo32.229.4
OmniScience Accuracy30.818.8
OmniScience Non-Hallucination17.349.5
QwenClawBench51.852.6
QwenWebBench11861397
RealWorldQA83.985.3
scicode4235.8
SimpleVQA67.158.9
SkillsBench Avg53028.7
supergpqa70.464.7
SWE-bench Multilingual69.367.2
SWE-bench Pro50.949.5
SWE-bench Verified76.473.4
TauBench V3 - Banking13.49.3
Terminal-Bench 2.052.551.5
Terminal-Bench 2.151.344.9
Terminal-Bench Hard40.934.8
toolathlon38.326.9
VideoMMMU84.783.7
VITA-Bench49.735.6
WideSearch7460.1
τ²-Bench Telecom (AA run)95.695.3
τ³-Bench13.467.2

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.