VECTOR WIREAI INTELLIGENCE
NVDA$1,847+3.2%MSFT$512+1.1%GOOGL$199-0.4%META$728+2.7%AMD$184-1.2%TSM$212+0.6%PLTR$98+4.1%AI IDX4,821+1.9%
PKT
SEEDRefresh Models Deals Regulatory Sources

Claude 3.7 Sonnet vs GPT-4o mini

AnthropicvsOpenAI21 shared benchmarks182 head-to-head
BenchmarkClaude 3.7 SonnetGPT-4o mini
AA Agentic Index371
AA Intelligence27.67
aider_polyglot64.93.6
AIR-Bench 202481.856.3
anthropic_red_team99.798.3
ARC-AGI-200
Artificial Analysis Coding Index36.411.4
bbq92.188.2
Fortress3848.1
gdpval27.40
GPQA Diamond84.842.6
harmbench84.384.9
HLE9.74.2
IFBench48.331
MMMU7559.4
MMMU-Pro60.141.5
scicode40.322.9
simple_safety_tests10097.8
simplebench46.410.7
SWE-bench Verified70.38.7
xstest96.496

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.