VECTOR WIREAI INTELLIGENCE
UTC
Refresh Models Deals Regulatory Sources

Qwen3.8 Max Preview vs Qwen3 VL 4B (Reasoning)

19 SHARED BENCHMARKS

Across 19 shared benchmarks, Qwen3.8 Max Preview scores higher on 19 and Qwen3 VL 4B (Reasoning) on 0. The widest gap is HLE, where Qwen3.8 Max Preview scores 43.6 against 4.6.

ALIBABAVSALIBABA19 SHARED190 HEAD-TO-HEAD

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.

Compare Qwen3.8 Max Preview withEVERY TRACKED PAIRING

Anthropic12 PAIRINGS
OpenAI11 PAIRINGS
Google11 PAIRINGS
Alibaba10 PAIRINGS
DeepSeek5 PAIRINGS
Moonshot4 PAIRINGS
Meta3 PAIRINGS
Z.ai1 PAIRING
NVIDIA1 PAIRING

Compare Qwen3 VL 4B (Reasoning) withEVERY TRACKED PAIRING

Anthropic12 PAIRINGS
OpenAI11 PAIRINGS
Google11 PAIRINGS
Alibaba10 PAIRINGS
DeepSeek5 PAIRINGS
Moonshot4 PAIRINGS
Meta3 PAIRINGS
Z.ai1 PAIRING
NVIDIA1 PAIRING