VECTOR WIREAI INTELLIGENCE
NVDA$1,847+3.2%MSFT$512+1.1%GOOGL$199-0.4%META$728+2.7%AMD$184-1.2%TSM$212+0.6%PLTR$98+4.1%AI IDX4,821+1.9%
PKT
SEEDRefresh Models Deals Regulatory Sources

Claude Sonnet 5 vs Qwen3.6 Plus

AnthropicvsAlibaba33 shared benchmarks285 head-to-head
BenchmarkClaude Sonnet 5Qwen3.6 Plus
AA Agentic Index49.729
AA Intelligence55.340.5
AA-LCR7772.3
AA-Omniscience16.42.6
arena_elo14631444
arena_text_factuality14531438
Artificial Analysis Coding Index71.554.5
coding_arena_elo15401460
critpt16.92.9
gdpval54.832
GPQA Diamond91.190.4
HLE57.428.8
livebench_agentic_coding59.441.4
livebench_coding80.778.2
livebench_data_analysis71.769.9
livebench_instruction_following63.958.3
livebench_language7575
livebench_math92.983.7
livebench_reasoning88.775.8
MMMU-Pro77.378.8
OmniScience Accuracy4026.4
OmniScience Non-Hallucination60.668
OSWorld-Verified81.262.5
scicode53.640.7
simpleqa_verified2549.1
SWE-bench Multilingual78.373.8
SWE-bench Pro63.256.6
SWE-bench Verified85.278.8
TauBench V3 - Banking37.320.8
Terminal-Bench 2.080.461.6
Terminal-Bench 2.180.561.4
toolathlon54.339.8
τ³-Bench28.270.7

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.