VECTOR WIREAI INTELLIGENCE
UTC
Refresh Models Deals Regulatory Sources

Llama 3.1 Instruct 405B vs Qwen3.5 2B

25 SHARED BENCHMARKS

Across 25 shared benchmarks, Llama 3.1 Instruct 405B scores higher on 15 and Qwen3.5 2B on 8, with 2 level. The widest gap is AA Agentic Index, where Llama 3.1 Instruct 405B scores 6.3 against 0.8.

METAVSALIBABA25 SHARED158 HEAD-TO-HEAD

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.

Compare Llama 3.1 Instruct 405B withALL PAIRINGS →

Compare Qwen3.5 2B withALL PAIRINGS →