Command A vs Jamba 1.7 Mini
Wins 5 of 6 areas
Coding · Reasoning · Facts · Long documents · Following instructions
Wins 0 of 6 areas
—
Command A is the stronger all-rounder.
Scores updated · 19 tests both models report · How we compare
Where each one wins
Tests won in each of the six areas where both have results. Each piece is one test, so a longer bar means more evidence; grey means the two scored within a point of each other.
- FactsGetting facts right instead of making them up30Command A3 of 3 tests
- ReasoningHard problems that need careful thinking10Command A1 of 3 tests · 2 ties
- CodingWriting and fixing software10Command A1 of 2 tests · 1 tie
- Long documentsFinding answers in very long texts10Command A1 of 1 test
- Following instructionsDoing exactly what it is asked10Command A1 of 1 test
- AgentsCarrying out multi-step tasks on its own00Even0 each · 1 tie
Agents, long documents and following instructions rest on a single test each.
What it costs
Prices per million tokens, roughly 750,000 words. The bars show the cost of a million tokens read plus a million written.
The biggest differences
The tests each model wins by the widest margin, up to three each. Scores are out of 100.
Where Command A pulls ahead
- Graduate-level biology, physics and chemistry questionsGPQA Diamond+20.5points ahead
- Code for real scientific research problemsSciCode+18.8points ahead
- Avoids making up answers it doesn't knowAA-Omniscience · Non-hallucination+17.8points ahead
Where Jamba 1.7 Mini pulls ahead
No clear win on a test scored out of 100.
Every test, side by side
All 19 tests both models report. The winning score is in its model's colour; marks a score checked independently.
ReasoningCommand A
- GPQA DiamondCommand A by 20.552.732.2+20.5
- Humanity's Last Examtie44.4tie
- CritPttie00tie
FactsCommand A
- AA-Omniscience · Non-hallucinationCommand A by 17.822.74.9+17.8
- Vectara HHEM hallucination ratelower is betterCommand A by 5.49.314.7+5.4
- AA-Omniscience · AccuracyCommand A by 5.116.411.3+5.1
Following instructionsCommand A
- IFBenchCommand A by 5.136.531.4+5.1
Other results8 tests, not counted
Tests outside the eight areas. They are not counted above: several are summary scores built from other tests, or the same test under another name.
- AA-OmniscienceCommand A by 24.8-48.3-73.1+24.8
- Artificial Analysis Coding IndexCommand A by 6.89.93.1+6.8
- vectara_factual_consistencyCommand A by 5.490.785.3+5.4
- vectara_avg_summary_lengthJamba 1.7 Mini by 34.7 rating points101.7136.4+34.7 rating
- τ²-Bench Telecom (AA run)Command A by 2.615.212.6+2.6
- AA IntelligenceCommand A by 1.875.2+1.8
- vectara_answer_rateJamba 1.7 Mini by 1.597.699.1+1.5
- AA Agentic Indextie5.14.2tie
Questions people ask
Which is better, Command A or Jamba 1.7 Mini?
Command A wins five of the six areas where both have results: coding, reasoning, facts, long documents and following instructions. Jamba 1.7 Mini wins none. They are level on agents.
Which is better for coding?
Command A. It wins 1 of the 2 coding tests both models report; Jamba 1.7 Mini wins none, and 1 is a tie.
How do you compare the two?
We use the 19 benchmark tests both models have published scores on. The verdict counts the 11 tests in the eight capability areas, and a gap under one point (ten on rating-style scales) counts as a tie. The other 8 are listed but not counted, because several are summary scores or repeat a test. Each score is the one shown on the model's own page.