Qwen3.5 27B vs Seed 1.8
27 SHARED BENCHMARKSAcross 27 shared benchmarks, Qwen3.5 27B scores higher on 16 and Seed 1.8 on 11. The widest gap is DynaMath, where Qwen3.5 27B scores 87.7 against 61.5. Seed 1.8 is the cheaper of the two on tracked API pricing ($0.25 against $0.30 per million input tokens).
At a glance
Qwen3.5 27B's release date per Artificial Analysis. Prices: Alibaba's own price page for Qwen3.5 27B; deepinfra for Seed 1.8. ◆ = independently verified score.
Where each leadsby capability area · benchmark wins
Biggest gaps
Qwen3.5 27B pulls furthest ahead on
- CharXiv (RQ)79.5 vs 71.4
Seed 1.8 pulls furthest ahead on
- BrowseComp67.6 vs 61
- OSWorld-Verified61.9 vs 56.2
- Multi-Challenge66.7 vs 60.8
Every shared benchmark27 · grouped by area
Coding 2
Agentic 2
Instruction Following 1
Multimodal 5
Other shared benchmarks 17
Each score is the one the model's own page shows — the most authoritative tracked result for that benchmark, on the benchmark's standard methodology (source-attributed). ↓ marks lower-is-better metrics; ◆ an independently verified score. Open either model for its full surface, provenance and pricing. Quoted rates are the price-setter row we currently track for each model — its direct or vendor-official listing where one exists (alibaba-official, deepinfra), otherwise the lowest tracked offer.