VECTOR WIREAI INTELLIGENCE
PKT
Refresh Models Deals Regulatory Sources

Questions this page answers

Which is better, Claude Opus 4.6 or muse-glimmer-30b?
Across 17 shared benchmarks, Claude Opus 4.6 scores higher on 12 and muse-glimmer-30b on 5. The widest gap is baby_vision, where muse-glimmer-30b scores 70.4 against 14.8. muse-glimmer-30b is the cheaper of the two on tracked API pricing ($0.30 against $5.00 per million input tokens).

Claude Opus 4.6 vs muse-glimmer-30b

Across 17 shared benchmarks, Claude Opus 4.6 scores higher on 12 and muse-glimmer-30b on 5. The widest gap is baby_vision, where muse-glimmer-30b scores 70.4 against 14.8. muse-glimmer-30b is the cheaper of the two on tracked API pricing ($0.30 against $5.00 per million input tokens).

AnthropicvsMeta17 shared benchmarks125 head-to-head
BenchmarkClaude Opus 4.6muse-glimmer-30b
AA-LCR74.380
aime_202696.794.7
baby_vision14.870.4
charxiv_rq69.185.9
deepsearchqa_f191.374.6
GPQA Diamond91.383.5
HLE53.122
IFBench62.577
MCP Atlas76.875.5
MMMU-Pro77.374
OmniDocBench 1.586.675.8
OSWorld-Verified72.765.9
scicode5243.6
screenspot_pro_no_tools57.775.4
SWE-bench Pro57.351.2
SWE-bench Verified80.876
Terminal Bench 2.1 (Terminus)78.251.7

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing. Quoted rates are the price-setter row we currently track for each model — its direct or vendor-official listing where one exists (anthropic-official, deepinfra), otherwise the lowest tracked offer.