VECTOR WIREAI INTELLIGENCE
PKT
Refresh Models Deals Regulatory Sources

Questions this page answers

Which is better, Kimi K2.6 or muse-glimmer-30b?
Across 16 shared benchmarks, Kimi K2.6 scores higher on 11 and muse-glimmer-30b on 5. The widest gap is HLE, where Kimi K2.6 scores 37.5 against 22. muse-glimmer-30b is the cheaper of the two on tracked API pricing ($0.30 against $0.91 per million input tokens).

Kimi K2.6 vs muse-glimmer-30b

Across 16 shared benchmarks, Kimi K2.6 scores higher on 11 and muse-glimmer-30b on 5. The widest gap is HLE, where Kimi K2.6 scores 37.5 against 22. muse-glimmer-30b is the cheaper of the two on tracked API pricing ($0.30 against $0.91 per million input tokens).

MoonshotvsMeta16 shared benchmarks115 head-to-head
BenchmarkKimi K2.6muse-glimmer-30b
AA-LCR76.780
aime_202696.494.7
baby_vision68.570.4
charxiv_rq86.785.9
deepsearchqa_f192.574.6
GPQA Diamond91.183.5
HLE37.522
IFBench7677
MCP Atlas68.175.5
MMMU-Pro80.174
OSWorld-Verified73.165.9
scicode53.543.6
SWE-bench Pro58.651.2
SWE-bench Verified80.276
Tau 3 Banking20.623.5
Terminal-Bench 2.165.951.7

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing. Quoted rates are the price-setter row we currently track for each model — its direct or vendor-official listing where one exists (moonshot-official, deepinfra), otherwise the lowest tracked offer.