Qwen2.5 Instruct 72B
Qwen2.5 Instruct 72B has too few ranked results yet to rate it on any capability.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability.
No capability has enough ranked results to rate yet.
Price
$0.47input$0.49outputper million tokens
From Artificial Analysis · 4 providers tracked · All prices
Evidence
144results on106benchmarks
- 13 independently verified
- 13 aggregator
- 45 vendor-reported
- 73 cross-referenced
From 19 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingStructured outputs
As listed by OpenRouter
Qwen2.5 Instruct 72B benchmark results
144 results on 106 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
0 of 6 ranked benchmarks measured
- 0.00Oct 8, 2026
- GPQA Diamond49.09Oct 8, 2026gpqa
- Humanity's Last Exam3.56Oct 8, 2026aa_hle
Show 2 more reasoning resultsHide 2 reasoning results
- GPQA Diamond49.00Oct 7, 2026GPQA
- GPQA Diamond49.00Jun 6, 2026GPQA (diamond)
0 of 10 ranked benchmarks measured
- Terminal-Bench Hard4.55Oct 8, 2026aa_terminalbench_hard
- SciCode26.74Sep 4, 2026aa_scicode
- 23.80May 3, 2026
0 of 3 ranked benchmarks measured
- 20.00Sep 4, 2026
- 39.40May 3, 2026
0 of 3 ranked benchmarks measured
- IFBench36.87Oct 8, 2026aa_ifbench
0 of 4 ranked benchmarks measured
- AA-Omniscience · Accuracy17.47Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination14.54Oct 8, 2026omniscienceNonHallucination
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- AA Intelligence10.00Jul 3, 2026Artificial Analysis Intelligence Index
- 96.20Jul 2, 2026
- 90.00Jul 2, 2026
- 58.97May 19, 2026
- 97.94May 10, 2026
- 100.00May 10, 2026
Show 125 more resultsHide 125 results
- AA Intelligence7.73Oct 8, 2026aa_intelligence_index
- -53.07Oct 8, 2026
- AGIEval75.80Aug 31, 2026AGIEval (Acc.)
- AGIEval75.80May 1, 2026AGIEval (Acc.)
- 65.40May 3, 2026
- 7.60May 3, 2026
- 23.30May 3, 2026
- 81.60Oct 7, 2026
- 99.56May 10, 2026
- 94.50May 1, 2026
- 94.50Jul 6, 2026
- 98.40May 1, 2026
- 81.20Sep 8, 2026
- 81.20Jun 6, 2026
- 11.94Jun 18, 2026
- 79.80May 1, 2026
- 79.80Jul 6, 2026
- 95.40May 10, 2026
- C-Eval89.20Aug 31, 2026C-Eval (Acc.)
- 86.10May 3, 2026
- C-Eval89.20May 1, 2026C-Eval (Acc.)
- 76.70May 1, 2026
- 76.70Jul 6, 2026
- CCPM88.50Aug 31, 2026CCPM (Acc.)
- CCPM88.50May 1, 2026CCPM (Acc.)
- Chinese SimpleQA (C-SimpleQA)52.20Jun 6, 2026C-SimpleQA
- 48.40May 3, 2026
- CLUEWSC82.50Aug 31, 2026CLUEWSC (EM)
- 91.40May 3, 2026
- CLUEWSC82.50May 1, 2026CLUEWSC (EM)
- 84.50May 1, 2026
- CMMLU89.50May 1, 2026chinese_cmmlu_acc
- 89.50Jul 6, 2026
- CMRC75.80Aug 31, 2026CMRC (EM)
- CMRC75.80May 1, 2026CMRC (EM)
- 15.90May 3, 2026
- 24.80May 3, 2026
- CRUXEval-I (input prediction)59.10May 1, 2026CRUXEval-I (Acc.)
- CRUXEval-O (output prediction)59.90May 1, 2026CRUXEval-O (Acc.)
- 76.70May 30, 2026
- 76.70May 3, 2026
- 80.60May 1, 2026
- 85.00Jun 6, 2026
- 69.80May 3, 2026
- 95.80Oct 7, 2026
- 95.80Jun 6, 2026
- 88.30May 1, 2026
- 72.81May 10, 2026
- 84.80May 1, 2026
- 84.80Jul 6, 2026
- 86.60Oct 7, 2026
- HumanEval53.00May 1, 2026HumanEval (Pass@1)
- 86.60Jun 6, 2026
- 80.40May 30, 2026
- 77.30May 3, 2026
- 84.10Aug 31, 2026
- 87.20Jun 12, 2026
- 84.10May 3, 2026
- 87.20Jun 6, 2026
- 52.30Oct 7, 2026
- 55.50Aug 23, 2026
- LiveCodeBench28.70May 3, 2026LiveCodeBench (Pass@1)
- LiveCodeBench12.90May 1, 2026code_livecodebenchbase_pass1
- 31.10May 3, 2026
- 12.90Jul 6, 2026
- 47.90Jun 6, 2026
- 42.70Jun 6, 2026
- 40.80Jun 6, 2026
- 41.80Jun 6, 2026
- 39.80Jun 12, 2026
- 44.40Jun 6, 2026
- 40.90Jun 6, 2026
- 38.10Jun 6, 2026
- 43.50Jun 6, 2026
- 42.10Jun 6, 2026
- 48.90Jun 6, 2026
- 45.60Jun 6, 2026
- 88.39May 10, 2026
- 81.80Jun 6, 2026
- 80.00May 30, 2026
- 54.40May 1, 2026
- 54.40Jul 6, 2026
- 80.00May 3, 2026
- 88.20Oct 7, 2026
- MBPP72.60May 1, 2026MBPP (Pass@1)
- 77.00Jun 6, 2026
- 87.30May 30, 2026
- 76.20May 1, 2026
- 76.20Jul 6, 2026
- 76.96May 10, 2026
- 86.10Jun 6, 2026
- 85.30May 30, 2026
- 83.20May 1, 2026
- 85.00Jul 6, 2026
- 71.10Oct 7, 2026
- MMLU-Pro58.30Jul 3, 2026english_mmlupro_acc
- 71.10Jun 6, 2026
- 71.60May 3, 2026
- 58.30Jul 6, 2026
- 86.80Oct 7, 2026
- 85.60May 3, 2026
- 83.20Jul 6, 2026
- MMMLU74.80Jul 6, 2026Multilingual (MMMLU-non-English (Acc.))
- 74.80Jul 7, 2026
- 9.35Oct 7, 2026
- 75.10Oct 7, 2026
- 74.47May 10, 2026
- 33.20May 1, 2026
- 33.20Jul 6, 2026
- 35.94May 10, 2026
- Pile-test (BPB)lower is better0.64Jul 6, 2026
- 82.60May 1, 2026
- 82.60Jul 6, 2026
- RACE-High50.30Aug 31, 2026RACE-High (Acc.)
- RACE-High50.30May 1, 2026RACE-High (Acc.)
- RACE-Middle68.10Aug 31, 2026RACE-Middle (Acc.)
- RACE-Middle68.10May 1, 2026RACE-Middle (Acc.)
- 10.30Jun 6, 2026
- 10.20May 30, 2026
- 9.10May 3, 2026
- The Pile (Test, BPB)lower is better0.64May 1, 2026
- 71.90May 1, 2026
- 82.30May 1, 2026
- 82.30Jul 6, 2026
- τ²-Bench Telecom (AA run)34.50Oct 8, 2026aa_tau2
Qwen2.5 Instruct 72B: common questions
Who makes Qwen2.5 Instruct 72B?
Qwen2.5 Instruct 72B is made by Alibaba.
When was Qwen2.5 Instruct 72B released?
Qwen2.5 Instruct 72B was released on Sep 19, 2024, according to Artificial Analysis.
How much does Qwen2.5 Instruct 72B cost?
Qwen2.5 Instruct 72B costs $0.47 per million input tokens and $0.49 per million output tokens, according to Artificial Analysis. We track its price at 4 providers. At a mix of three input tokens to one output token, it is cheaper than 60% of the 329 priced models we track.
How many benchmarks has Qwen2.5 Instruct 72B been tested on?
We track 144 results for Qwen2.5 Instruct 72B on 106 benchmarks from 19 sources, 13 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does Qwen2.5 Instruct 72B support?
OpenRouter lists tool calling and structured outputs for Qwen2.5 Instruct 72B.
About this record
Where Qwen2.5 Instruct 72B's numbers come from, and every name it appears under.
- Tracked since
- May 2, 2026
- Newest source mention
- Sep 2, 2026
Where the results come from
Verification: 144 scores · 13 independently verified · 13 aggregator-attributed · 73 vendor cross-reference · 45 vendor-reported. How these tiers are assigned
From 19 sources on 6 sites. raw.githubusercontent.com supplies 54 of them; the 13 independently verified results come from 2 sites. Bars are coloured by trust tier.
- raw.githubusercontent.com54
- huggingface.co31
- arxiv.org20
- artificialanalysis.ai14
- api.llm-stats.com13
- storage.googleapis.com12