Across 44 shared benchmarks, Qwen3.5 122B A10B scores higher on 40 and Qwen3 VL 235B A22B Reasoning on 4. The widest gap is HLE, where Qwen3.5 122B A10B scores 47.5 against 13.6. Both cost $0.40 per million input tokens on tracked API pricing; Qwen3.5 122B A10B is cheaper on output ($3.20 against $4.00).
Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing. Quoted rates are the price-setter row we currently track for each model — its direct or vendor-official listing where one exists (direct, alibaba-official), otherwise the lowest tracked offer.