VECTOR WIREAI INTELLIGENCE
PKT
Refresh Models Deals Regulatory Sources

Questions this page answers

Which is better, Kimi K2 (Reasoning) or Qwen3.5 122B A10B?
Across 30 shared benchmarks, Kimi K2 (Reasoning) scores higher on 12 and Qwen3.5 122B A10B on 16, with 2 level. The widest gap is critpt, where Kimi K2 (Reasoning) scores 2.6 against 0.9. Tracked API pricing per million tokens: Kimi K2 (Reasoning) $0.60 in / $2.50 out, Qwen3.5 122B A10B $0.40 in / $3.20 out.

Kimi K2 (Reasoning) vs Qwen3.5 122B A10B

Across 30 shared benchmarks, Kimi K2 (Reasoning) scores higher on 12 and Qwen3.5 122B A10B on 16, with 2 level. The widest gap is critpt, where Kimi K2 (Reasoning) scores 2.6 against 0.9. Tracked API pricing per million tokens: Kimi K2 (Reasoning) $0.60 in / $2.50 out, Qwen3.5 122B A10B $0.40 in / $3.20 out.

MoonshotvsAlibaba30 shared benchmarks1216 head-to-head
BenchmarkKimi K2 (Reasoning)Qwen3.5 122B A10B
AA Agentic Index47.921.3
AA Intelligence33.532.8
AA-LCR70.370.3
AA-Omniscience-21.4-41.5
Artificial Analysis Coding Index34.845.7
browsecomp60.263.8
browsecomp_zh62.369.9
critpt2.60.9
gdpval24.524.3
GPQA Diamond84.586.6
HLE23.947.5
HMMT 202589.490.3
IFBench68.176.1
LiveCodeBench v683.178.9
longbench_v245.160.2
mmlu_redux94.494
MMLU-Pro84.686.7
multichallenge66.461.5
OmniScience Accuracy30.924.4
OmniScience Non-Hallucination25.812.9
scicode44.842
Seal-056.344.1
SWE-bench Verified71.372
Terminal-Bench 2.035.749.4
Terminal-Bench Hard31.131.1
vectara_answer_rate98.699.8
vectara_avg_summary_length59.286.4
vectara_factual_consistency82.188.8
vectara_hallucination_rate17.911.2
τ²-Bench Telecom (AA run)9393.6

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing. Quoted rates are the price-setter row we currently track for each model — its direct or vendor-official listing where one exists (direct, direct), otherwise the lowest tracked offer.