VECTOR WIREAI INTELLIGENCE
UTC
Refresh Models Deals Regulatory Sources

Llama 3.1 Instruct 405B vs Qwen3.5 4B

22 SHARED BENCHMARKS

Across 22 shared benchmarks, Llama 3.1 Instruct 405B scores higher on 5 and Qwen3.5 4B on 16, with 1 level. The widest gap is AA Agentic Index, where Qwen3.5 4B scores 36.3 against 6.3.

METAVSALIBABA22 SHARED516 HEAD-TO-HEAD

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.

Compare Llama 3.1 Instruct 405B withALL PAIRINGS →

Compare Qwen3.5 4B withALL PAIRINGS →