VECTOR WIREAI INTELLIGENCE
UTC
Refresh Models Deals Regulatory Sources

Kimi K2 Instruct vs Llama 3.1 Instruct 405B

34 SHARED BENCHMARKS

Across 34 shared benchmarks, Kimi K2 Instruct scores higher on 31 and Llama 3.1 Instruct 405B on 2, with 1 level. The widest gap is AA Agentic Index, where Kimi K2 Instruct scores 37.7 against 6.3.

MOONSHOTVSMETA34 SHARED312 HEAD-TO-HEAD

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.

Compare Kimi K2 Instruct withEVERY TRACKED PAIRING

Anthropic12 PAIRINGS
OpenAI11 PAIRINGS
Google11 PAIRINGS
Alibaba11 PAIRINGS
DeepSeek5 PAIRINGS
Moonshot3 PAIRINGS
Meta1 PAIRING
Z.ai2 PAIRINGS
MiniMax1 PAIRING
NVIDIA1 PAIRING

Compare Llama 3.1 Instruct 405B withEVERY TRACKED PAIRING

Anthropic12 PAIRINGS
OpenAI11 PAIRINGS
Google11 PAIRINGS
Alibaba11 PAIRINGS
DeepSeek5 PAIRINGS
Moonshot3 PAIRINGS
Meta1 PAIRING
Z.ai2 PAIRINGS
MiniMax1 PAIRING
NVIDIA1 PAIRING