Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

Claude Opus 4.8 vs Gemini 4 Argon

7 SHARED BENCHMARKS

Across 7 shared benchmarks, Claude Opus 4.8 scores higher on 0 and Gemini 4 Argon on 7. The widest gap is Terminal-Bench 4.0, where Gemini 4 Argon scores 57.4 against 23.6. Gemini 4 Argon is the cheaper of the two on tracked API pricing ($2.00 against $5.00 per million input tokens).

ANTHROPICVSGOOGLE7 SHARED0–7 HEAD-TO-HEADNEWEST SCORE ADDED

At a glance

MakerAnthropicGoogle
Released—
Price per 1M tokens input / output$5.00 / $25.00$2.00 / $10.00
Cost of 1M in + 1M out$30.00$12.00 2.5× less
Head-to-head of 7 shared benchmarks0 wins7 wins
Scores tracked independently verified200 32 ◆13 1 ◆

Claude Opus 4.8's release date per Artificial Analysis. Prices: Anthropic's own price page for Claude Opus 4.8; Artificial Analysis for Gemini 4 Argon. ◆ = independently verified score.

Every shared benchmark7 · grouped by area

Other shared benchmarks 7

AA Intelligence41.8◆ 53
OSWorld 2.055.769.2
Terminal-Bench 4.023.657.4

Each score is the one the model's own page shows — the most authoritative tracked result for that benchmark, on the benchmark's standard methodology (source-attributed). ↓ marks lower-is-better metrics; ◆ an independently verified score. Open either model for its full surface, provenance and pricing. Quoted rates are the price-setter row we currently track for each model — its direct or vendor-official listing where one exists (anthropic-official, direct), otherwise the lowest tracked offer.

Compare Claude Opus 4.8 withALL PAIRINGS →