VECTOR WIREAI INTELLIGENCE
PKT
Refresh Models Deals Regulatory Sources

Questions this page answers

Which is better, Kimi K2.5 or muse-glimmer-30b?
Across 17 shared benchmarks, Kimi K2.5 scores higher on 9 and muse-glimmer-30b on 8. The widest gap is HLE, where Kimi K2.5 scores 50.2 against 22. muse-glimmer-30b is the cheaper of the two on tracked API pricing ($0.30 against $0.56 per million input tokens).

Kimi K2.5 vs muse-glimmer-30b

Across 17 shared benchmarks, Kimi K2.5 scores higher on 9 and muse-glimmer-30b on 8. The widest gap is HLE, where Kimi K2.5 scores 50.2 against 22. muse-glimmer-30b is the cheaper of the two on tracked API pricing ($0.30 against $0.56 per million input tokens).

MoonshotvsMeta17 shared benchmarks98 head-to-head
BenchmarkKimi K2.5muse-glimmer-30b
AA-LCR7380
aime_202695.894.7
baby_vision36.570.4
charxiv_rq78.785.9
deepsearchqa_f18974.6
GPQA Diamond87.983.5
HLE50.222
IFBench70.277
MCP Atlas6475.5
MMMU-Pro78.574
OmniDocBench 1.588.875.8
OSWorld-Verified63.365.9
scicode4943.6
SWE-bench Pro53.851.2
SWE-bench Verified76.876
Tau 3 Banking14.223.5
Terminal-Bench 2.145.751.7

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing. Quoted rates are the price-setter row we currently track for each model — its direct or vendor-official listing where one exists (moonshot-official, deepinfra), otherwise the lowest tracked offer.