VECTOR WIREAI INTELLIGENCE
NVDA$1,847+3.2%MSFT$512+1.1%GOOGL$199-0.4%META$728+2.7%AMD$184-1.2%TSM$212+0.6%PLTR$98+4.1%AI IDX4,821+1.9%
PKT
SEEDRefresh Models Deals Regulatory Sources

Claude Fable 5 vs Claude Opus 4.6

AnthropicvsAnthropic52 shared benchmarks457 head-to-head
BenchmarkClaude Fable 5Claude Opus 4.6
AA Agentic Index56.667.6
AA Intelligence62.144.9
AA-LCR76.774.3
AA-Omniscience43.313.7
arc_agi_190.593
ARC-AGI-276.868.8
arena_elo15081497
arena_text_factuality14861487
arena_vision13071300
Artificial Analysis Coding Index76.548.1
baby_vision_with_python90.538.4
browsecomp8886.8
charxiv_rq89.469.1
coding_arena_elo16261545
critpt28.612.6
deepsearchqa_f194.291.3
gdpval61.955.9
GDPval-AA (Elo)19321619
GPQA Diamond92.691.3
HiL-Bench56.338.3
HLE64.553.1
HLE (with tools)64.562.7
IFBench63.562.5
Legal Agent Benchmark13.34.2
livebench78.376.3
livebench_agentic_coding62.249
livebench_coding8678.2
livebench_data_analysis80.569.9
livebench_instruction_following75.863.3
livebench_language90.783.3
livebench_math9689.3
livebench_reasoning89.788.7
mathvision94.871.2
MCP Atlas84.776.8
MMMU-Pro84.277.3
officeqa_pro69.957.1
OmniScience Accuracy65.347
OmniScience Non-Hallucination36.437.2
OSWorld-Verified8572.7
scicode60.252
simplebench81.967.6
simpleqa_verified68.346.5
swe_bench_multimodal54.127.1
SWE-bench Multilingual86.677.8
SWE-bench Pro8057.3
SWE-bench Verified9580.8
Terminal-Bench Hard62.965.4
toolathlon61.756.8
Toolathlon Avg turns19.816.9
Toolathlon Pass@∞55.647.2
τ²-Bench Telecom (AA run)98.599.3
τ³-Bench26.872.4

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.