VECTOR WIREAI INTELLIGENCE
NVDA$1,847+3.2%MSFT$512+1.1%GOOGL$199-0.4%META$728+2.7%AMD$184-1.2%TSM$212+0.6%PLTR$98+4.1%AI IDX4,821+1.9%
PKT
SEEDRefresh Models Deals Regulatory Sources

GPT-5.6 Sol vs Opus 5

OpenAIvsAnthropic44 shared benchmarks1528 head-to-head
BenchmarkGPT-5.6 SolOpus 5
AA Agentic Index57.859.2
AA Intelligence6163.1
AA-LCR77.778.7
AA-Omniscience2237.1
arc_agi_197.597.5
arc_agi_37.830.2
ARC-AGI-292.590.4
arena_elo14821487
arena_text_factuality14741481
Artificial Analysis Coding Index78.378
browsecomp90.490.8
coding_arena_elo16191662
critpt32.329.1
DeepSWE 1.17368.8
Finance Agent v253.858.6
GDP (Surge AI)30.785.5
gdpval61.167.2
GDPval-AA v217481861
GPQA Diamond94.693.7
HealthBench5767.1
healthbench_professional60.559.8
HiL-Bench32.357
HLE49.564.7
HLE (with tools)5864.7
livebench_agentic_coding56.265.2
livebench_coding83.981.5
livebench_data_analysis79.874.5
livebench_instruction_following71.863.8
livebench_language87.788.7
livebench_math96.295.7
livebench_reasoning91.791.2
MCP Atlas83.685.8
MMMU-Pro84.684.7
officeqa_pro63.266.9
OmniScience Accuracy59.460.9
OmniScience Non-Hallucination10.640.5
OSWorld 2.062.670.6
OSWorld-Verified834
scicode56.955.7
simplebench64.880.6
simpleqa_verified71.656.7
SWE-bench Pro64.679.2
TauBench V3 - Banking44.344.7
Terminal-Bench 2.189.589.1

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.