VECTOR WIREAI INTELLIGENCE
NVDA$1,847+3.2%MSFT$512+1.1%GOOGL$199-0.4%META$728+2.7%AMD$184-1.2%TSM$212+0.6%PLTR$98+4.1%AI IDX4,821+1.9%
PKT
SEEDRefresh Models Deals Regulatory Sources

Qwen3.5 27B vs Qwen3.6 Plus

AlibabavsAlibaba57 shared benchmarks749 head-to-head
BenchmarkQwen3.5 27BQwen3.6 Plus
AA Agentic Index54.629
AA Intelligence34.640.5
AA-LCR72.372.3
AA-Omniscience-442.6
AI2D92.994.4
aime_202692.695.3
Artificial Analysis Coding Index34.954.5
C-Eval90.593.3
Claw Eval (pass@3)46.258.7
coding_arena_elo13571460
critpt0.92.9
DeepPlanning22.641.5
ERQA60.565.7
gdpval3332
GPQA Diamond85.890.4
HLE48.528.8
HMMT 202589.894.6
hmmt_feb_202684.387.8
hmmt_nov_202589.894.6
IFBench76.575.2
ifeval9594.3
imo_answer_bench79.983.8
include81.685.1
LiveCodeBench v680.787.1
longbench_v260.662
mathvision8688
MCP Atlas68.474.1
mcpmark36.348.2
MLVU85.986.7
mmlu_prox82.284.7
mmlu_redux93.294.5
MMLU-Pro86.188.5
mmmlu85.989.5
MMMU82.386
MMMU-Pro7578.8
MMStar8183.3
nl2repo27.337.9
OmniDocBench 1.588.991.2
OmniScience Accuracy20.726.4
OmniScience Non-Hallucination24.968
OSWorld-Verified56.262.5
RealWorldQA83.785.4
scicode39.540.7
screenspot_pro_no_tools70.368.2
SimpleVQA560.7
supergpqa65.671.6
SWE-bench Multilingual69.373.8
SWE-bench Pro51.256.6
SWE-bench Verified7578.8
Terminal-Bench 2.041.661.6
Terminal-Bench Hard32.643.9
tool_decathlon31.539.8
VideoMMMU82.384
VITA-Bench41.944.3
WideSearch66.474.3
τ²-Bench Telecom (AA run)93.997.7
τ³-Bench68.470.7

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.