VECTOR WIREAI INTELLIGENCE
PKT
Refresh Models Deals Regulatory Sources

Questions this page answers

Which model leads FLTEval pass@1?
Across 6 models scored on FLTEval pass@1, Claude Opus 4.6 leads at 39.6, ahead of Leanstral 1.5 at 28.9. The median tracked score is 23.7, and the field spans 16.6 to 39.6.

FLTEval pass@1

Across 6 models scored on FLTEval pass@1, Claude Opus 4.6 leads at 39.6, ahead of Leanstral 1.5 at 28.9. The median tracked score is 23.7, and the field spans 16.6 to 39.6.

6 models tracked
Data as of August 25, 2026
#ModelVendorBest scoreRunsLast seen
1Claude Opus 4.6Anthropic39.612026-08-25
2Leanstral 1.5Mistral28.912026-07-03
3Qwen3.5 397B A17BAlibaba25.412026-07-08
4Claude Sonnet 4.6Anthropic23.712026-08-25
5Claude Haiku 4.5Anthropic2312026-08-25
6GLM5-744B-A40BZ.ai16.612026-07-10

Best tracked score per model (default configuration; source-attributed and verification-tiered). Open a model for its full benchmark surface, provenance and pricing.