Qwen3 VL 235B A22B Reasoning vs Seed 1.8
22 SHARED BENCHMARKSAcross 22 shared benchmarks, Qwen3 VL 235B A22B Reasoning scores higher on 3 and Seed 1.8 on 19. The widest gap is ZeroBench, where Seed 1.8 scores 11 against 4. Seed 1.8 is the cheaper of the two on tracked API pricing ($0.25 against $0.40 per million input tokens).
At a glance
Qwen3 VL 235B A22B Reasoning's release date per Artificial Analysis. Prices: Alibaba's own price page for Qwen3 VL 235B A22B Reasoning; deepinfra for Seed 1.8. ◆ = independently verified score.
Where each leadsby capability area · benchmark wins
Biggest gaps
Qwen3 VL 235B A22B Reasoning pulls furthest ahead on
No ratified-area lead of 3 points or more.
Seed 1.8 pulls furthest ahead on
- OSWorld-Verified61.9 vs 38.1
- LiveCodeBench v679.5 vs 70.1
- BLINK74.3 vs 67.1
Every shared benchmark22 · grouped by area
Coding 1
Agentic 1
Multimodal 6
Other shared benchmarks 14
Each score is the one the model's own page shows — the most authoritative tracked result for that benchmark, on the benchmark's standard methodology (source-attributed). ↓ marks lower-is-better metrics; ◆ an independently verified score. Open either model for its full surface, provenance and pricing. Quoted rates are the price-setter row we currently track for each model — its direct or vendor-official listing where one exists (alibaba-official, deepinfra), otherwise the lowest tracked offer.