VECTOR WIREAI INTELLIGENCE
NVDA$1,847+3.2%MSFT$512+1.1%GOOGL$199-0.4%META$728+2.7%AMD$184-1.2%TSM$212+0.6%PLTR$98+4.1%AI IDX4,821+1.9%
PKT
SEEDRefresh Models Deals Regulatory Sources

Qwen3.5 397B A17B vs Qwen3.6 Plus

AlibabavsAlibaba65 shared benchmarks2439 head-to-head
BenchmarkQwen3.5 397B A17BQwen3.6 Plus
AA Agentic Index53.329
AA Intelligence34.340.5
AA-LCR72.772.3
AA-Omniscience-30.82.6
AgentWorldBench - Android54.957.6
AgentWorldBench - MCP68.355.3
AgentWorldBench - OS60.960.3
AgentWorldBench - Overall54.750.8
AgentWorldBench - Search30.821.9
AgentWorldBench - SWE64.459.1
AgentWorldBench - Terminal55.350.6
AgentWorldBench - Web48.550.8
aime_202694.295.3
arena_elo14421444
arena_text_factuality14451438
Artificial Analysis Coding Index48.254.5
C-Eval9393.3
Claw Eval (pass@3)48.158.7
coding_arena_elo13991460
critpt1.72.9
DeepPlanning34.341.5
ERQA67.565.7
gdpval35.832
GPQA Diamond89.390.4
HLE2928.8
HLE (with tools)48.350.6
HMMT 202592.794.6
hmmt_feb_202687.987.8
hmmt_nov_202592.794.6
IFBench78.875.2
ifeval92.694.3
imo_answer_bench80.983.8
include85.685.1
LiveCodeBench v683.687.1
longbench_v263.262
mcpmark46.148.2
MLVU86.786.7
mmlu_prox84.784.7
mmlu_redux94.994.5
MMLU-Pro87.888.5
mmmlu88.589.5
MMMU8586
MMMU-Pro7978.8
MMStar83.883.3
nl2repo32.237.9
OmniScience Accuracy30.826.4
OmniScience Non-Hallucination17.368
RealWorldQA83.985.4
scicode4240.7
simpleqa_verified2649.1
SimpleVQA67.10.7
supergpqa70.471.6
SWE-bench Multilingual69.373.8
SWE-bench Pro50.956.6
SWE-bench Verified76.478.8
TauBench V3 - Banking13.420.8
Terminal-Bench 2.052.561.6
Terminal-Bench 2.151.361.4
Terminal-Bench Hard40.943.9
toolathlon38.339.8
VideoMMMU84.784
VITA-Bench49.744.3
WideSearch7474.3
τ²-Bench Telecom (AA run)95.697.7
τ³-Bench13.470.7

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.