Across 32 shared benchmarks, Gemini 2.5 Flash scores higher on 17 and Kimi K2 Instruct on 15. The widest gap is τ²-Bench Telecom (AA run), where Kimi K2 Instruct scores 73.4 against 31.6. Tracked API pricing per million tokens: Gemini 2.5 Flash $0.30 in / $2.50 out, Kimi K2 Instruct $0.60 in / $2.50 out.
Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing. Quoted rates are the price-setter row we currently track for each model — its direct or vendor-official listing where one exists (google-official, direct), otherwise the lowest tracked offer.