VECTOR WIREAI INTELLIGENCE
PKT
Refresh Models Deals Regulatory Sources

Questions this page answers

Which is better, GPT-5 or Qwen3.5 122B A10B?
Across 39 shared benchmarks, GPT-5 scores higher on 27 and Qwen3.5 122B A10B on 12. The widest gap is critpt, where GPT-5 scores 5.7 against 0.9. Qwen3.5 122B A10B is the cheaper of the two on tracked API pricing ($0.40 against $1.25 per million input tokens).

GPT-5 vs Qwen3.5 122B A10B

Across 39 shared benchmarks, GPT-5 scores higher on 27 and Qwen3.5 122B A10B on 12. The widest gap is critpt, where GPT-5 scores 5.7 against 0.9. Qwen3.5 122B A10B is the cheaper of the two on tracked API pricing ($0.40 against $1.25 per million input tokens).

OpenAIvsAlibaba39 shared benchmarks2712 head-to-head
BenchmarkGPT-5Qwen3.5 122B A10B
AA Agentic Index49.721.3
AA Intelligence35.332.8
AA-LCR76.370.3
AA-Omniscience-8.7-41.5
arena_vision12121246
Artificial Analysis Coding Index38.945.7
browsecomp54.963.8
browsecomp_zh6569.9
coding_arena_elo14191358
critpt5.70.9
ERQA65.762
gdpval32.524.3
GPQA Diamond87.386.6
GSM8K-1.894.5
HLE28.547.5
HMMT 202593.390.3
IFBench73.176.1
LiveCodeBench v68778.9
mmlu_redux95.394
MMLU-Pro87.586.7
MMMU84.283.9
MMMU-Pro78.476.9
multichallenge71.161.5
OmniScience Accuracy40.324.4
OmniScience Non-Hallucination2112.9
scicode4342
Seal-051.444.1
SWE-bench Verified74.972
TauBench V3 - Banking22.115.3
Terminal-Bench 2.035.249.4
Terminal-Bench 2.135.247.6
Terminal-Bench Hard37.931.1
vectara_answer_rate99.999.8
vectara_avg_summary_length162.786.4
vectara_factual_consistency85.388.8
vectara_hallucination_rate14.711.2
VideoMMMU84.682
τ²-Bench Telecom (AA run)86.593.6
τ³-Bench19.613.6

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing. Quoted rates are the price-setter row we currently track for each model — its direct or vendor-official listing where one exists (direct, direct), otherwise the lowest tracked offer.