VECTOR WIREAI INTELLIGENCE
UTC
Refresh Models Deals Regulatory Sources

GPT-5.1 vs Kimi K2 Instruct

30 SHARED BENCHMARKS

Across 30 shared benchmarks, GPT-5.1 scores higher on 29 and Kimi K2 Instruct on 1. The widest gap is vectara_avg_summary_length, where GPT-5.1 scores 254.4 against 59.2. Kimi K2 Instruct is the cheaper of the two on tracked API pricing ($0.60 against $1.25 per million input tokens).

OPENAIVSMOONSHOT30 SHARED291 HEAD-TO-HEAD

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing. Quoted rates are the price-setter row we currently track for each model — its direct or vendor-official listing where one exists (direct, direct), otherwise the lowest tracked offer.

Compare GPT-5.1 withEVERY TRACKED PAIRING

Anthropic12 PAIRINGS
OpenAI10 PAIRINGS
Google11 PAIRINGS
Alibaba11 PAIRINGS
DeepSeek5 PAIRINGS
Moonshot3 PAIRINGS
Meta2 PAIRINGS
Z.ai2 PAIRINGS
MiniMax1 PAIRING
NVIDIA1 PAIRING

Compare Kimi K2 Instruct withEVERY TRACKED PAIRING

Anthropic12 PAIRINGS
OpenAI10 PAIRINGS
Google11 PAIRINGS
Alibaba11 PAIRINGS
DeepSeek5 PAIRINGS
Moonshot3 PAIRINGS
Meta2 PAIRINGS
Z.ai2 PAIRINGS
MiniMax1 PAIRING
NVIDIA1 PAIRING