VECTOR WIREAI INTELLIGENCE
PKT
Refresh Models Deals Regulatory Sources

Questions this page answers

Which is better, nvidia-nemotron-3-ultra-550b-a55b or Qwen 3.5 397B 17B?
Across 30 shared benchmarks, nvidia-nemotron-3-ultra-550b-a55b scores higher on 14 and Qwen 3.5 397B 17B on 16. The widest gap is Apex-Shortlist (with tools), where nvidia-nemotron-3-ultra-550b-a55b scores 84.8 against 60.4.

nvidia-nemotron-3-ultra-550b-a55b vs Qwen 3.5 397B 17B

Across 30 shared benchmarks, nvidia-nemotron-3-ultra-550b-a55b scores higher on 14 and Qwen 3.5 397B 17B on 16. The widest gap is Apex-Shortlist (with tools), where nvidia-nemotron-3-ultra-550b-a55b scores 84.8 against 60.4.

NVIDIAvsAlibaba30 shared benchmarks1416 head-to-head
Benchmarknvidia-nemotron-3-ultra-550b-a55bQwen 3.5 397B 17B
Airline81.576.5
Apex-Shortlist (no tools)74.961.4
Apex-Shortlist (with tools)84.860.4
Banking22.620.9
browsecomp44.440.5
CritPt (no tools)3.12.4
gdpval46.734.6
GPQA Diamond8787.1
HLE26.728.5
HLE (with tools)37.448.3
IFBench (prompt loose)81.778.2
imo_answer_bench92.384.5
IOI 2025570441.3
LiveCodeBench v68979.3
longbench_v261.968.9
MMLU-Pro86.888.3
MMLU-ProX (avg en/de/fr/es/it/ja/zh/hi/pt/ko)8386.4
multichallenge63.863.9
PinchBench9086.6
ProfBench (Search)5653
Retail86.488.5
SciCode (subtask)44.648
SWE-bench Multilingual67.770.9
SWE-bench Verified70.773.6
TauBench V3 - Average70.971
Telecom92.998
Terminal-Bench 2.156.449.9
Vals.ai Financial Agent 1.1 - with web search53.759
Vals.ai Financial Agent 1.1 - without web search60.161.3
WMT24++ (en→xx)83.786.8

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.