VECTOR WIREAI INTELLIGENCE
PKT
Refresh Models Deals Regulatory Sources

Questions this page answers

Which is better, MiniMax 2.7 230B A10B or nvidia-nemotron-3-ultra-550b-a55b?
Across 27 shared benchmarks, MiniMax 2.7 230B A10B scores higher on 4 and nvidia-nemotron-3-ultra-550b-a55b on 23. The widest gap is CritPt (no tools), where nvidia-nemotron-3-ultra-550b-a55b scores 3.1 against 0.6.

MiniMax 2.7 230B A10B vs nvidia-nemotron-3-ultra-550b-a55b

Across 27 shared benchmarks, MiniMax 2.7 230B A10B scores higher on 4 and nvidia-nemotron-3-ultra-550b-a55b on 23. The widest gap is CritPt (no tools), where nvidia-nemotron-3-ultra-550b-a55b scores 3.1 against 0.6.

MiniMaxvsNVIDIA27 shared benchmarks423 head-to-head
BenchmarkMiniMax 2.7 230B A10Bnvidia-nemotron-3-ultra-550b-a55b
Airline75.381.5
Apex-Shortlist (no tools)28.974.9
Apex-Shortlist (with tools)51.984.8
Banking14.622.6
browsecomp54.144.4
CritPt (no tools)0.63.1
gdpval47.646.7
GPQA Diamond86.687
HLE23.126.7
IFBench (prompt loose)74.681.7
imo_answer_bench75.192.3
LiveCodeBench v677.289
MMLU-Pro81.986.8
MMLU-ProX (avg en/de/fr/es/it/ja/zh/hi/pt/ko)78.483
multichallenge42.563.8
PinchBench77.690
ProfBench (Search)5256
Retail84.986.4
SciCode (subtask)38.344.6
SWE-bench Multilingual71.867.7
SWE-bench Verified75.370.7
TauBench V3 - Average66.170.9
Telecom89.692.9
Terminal-Bench 2.155.556.4
Vals.ai Financial Agent 1.1 - with web search50.553.7
Vals.ai Financial Agent 1.1 - without web search51.360.1
WMT24++ (en→xx)82.883.7

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.