Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

Claude Opus 5.5 vs Gemini 4 Argon

5 SHARED BENCHMARKS

Across 5 shared benchmarks, Claude Opus 5.5 scores higher on 4 and Gemini 4 Argon on 1. The widest gap is OSWorld 2.0, where Claude Opus 5.5 scores 81.8 against 69.2. Gemini 4 Argon is the cheaper of the two on tracked API pricing ($2.00 against $4.00 per million input tokens).

ANTHROPICVSGOOGLE5 SHARED4–1 HEAD-TO-HEADNEWEST SCORE ADDED

At a glance

MakerAnthropicGoogle
Released—
Price per 1M tokens input / output$4.00 / $20.00$2.00 / $10.00
Cost of 1M in + 1M out$24.00$12.00 2.0× less
Head-to-head of 5 shared benchmarks4 wins1 wins
Scores tracked independently verified101 28 ◆13 1 ◆

Claude Opus 5.5's release date per the vendor's own announcement. Prices: Anthropic's own price page for Claude Opus 5.5; Artificial Analysis for Gemini 4 Argon. ◆ = independently verified score.

Every shared benchmark5 · grouped by area

Other shared benchmarks 5

AA Intelligence57.6◆ 53
DeepSWE 1.174.277.9
OSWorld 2.081.869.2
Terminal-Bench 4.066.457.4
Terminal-Bench-Science 0.158.757.6

Each score is the one the model's own page shows — the most authoritative tracked result for that benchmark, on the benchmark's standard methodology (source-attributed). ↓ marks lower-is-better metrics; ◆ an independently verified score. Open either model for its full surface, provenance and pricing. Quoted rates are the price-setter row we currently track for each model — its direct or vendor-official listing where one exists (anthropic-official, direct), otherwise the lowest tracked offer.

Compare Claude Opus 5.5 withALL PAIRINGS →