Alibaba (Qwen), open weights

Qwen1.5-72B

Qwen1.5-72B by Alibaba (Qwen) ranks 285th of 354 ranked models on the Noometry Index as of October 2026, with a score of 30.8. Its strongest category is reasoning, where it ranks 203rd.

Last verified

Specifications

Noometry rank
#285 of 354
Index score
30.8
Evidence
Confirmed 22 results
Released
February 4, 2024
Weights
Open weights
Reasoning
Unknown
Context window
—
Max output
—
Input price
Not listed
Output price
Not listed
Blended price
Not listed
Output speed
Not measured
Value
Not ranked
Knowledge cutoff
Unknown

Category scores

Each category score combines every public result we have in that category.

Qwen1.5-72B category scores
  1. Coding 31.9
  2. Reasoning 22.2
  3. Math 33.2
  4. Knowledge 11.5
  5. Multilingual 33.2
  6. Instruction Following 59.3
  7. Long Context 35.1
  8. Writing & Preference 37.3
Qwen1.5-72B category ranks
CategoryScoreRankResults
Coding31.9#2773
Reasoning22.2#2031
Math33.2#2051
Knowledge11.5#3002
Multilingual33.2#2531
Instruction Following59.3#2561
Long Context35.1#2431
Writing & Preference37.3#2583

Strengths and weaknesses

Categories where Qwen1.5-72B places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

Qwen1.5-72B: strongest categories
CategoryScorevs medianRank
Reasoning22.2−1.4#203 of 350, top 58%
Math33.2−3.4#205 of 327, top 63%
Coding31.9−6.8#277 of 340, top 82%

Weakest categories

Qwen1.5-72B: weakest categories
CategoryScorevs medianRank
Knowledge11.5−25.8#300 of 314, top 96%
Multilingual33.2−14.2#253 of 297, top 86%
Instruction Following59.3−12.0#256 of 305, top 84%

Closest competitors

The models ranked just above and below Qwen1.5-72B. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Qwen1.5-72B
ModelRankScoreBlended $/MSpeed
Amazon Nova Pro#28131.0$1.40—Compare
Llama 4 Maverick#28230.9$0.30456Compare
Phi-4 Mini#28330.9$0.13—Compare
Gemma 3 27B#28430.8$0.1062Compare
Granite 3.0 2b Instruct#28630.8——Compare
Codellama 34b Instruct#28730.8——Compare
Llama 3.1-405B#28830.7—78Compare
Yi-1.5-34B#28930.6——Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

Qwen1.5-72B Coding benchmark results
BenchmarkScorePositionSettingSourceDate
BigCodeBench Instruct33.2%#50 of 64, top 79%BigCodeBench2024-04-26
LMArena Coding1165#253 of 294, top 87%LMArena2026-10-08
BigCodeBench Complete40.3%#53 of 66, top 81%BigCodeBench2024-04-26
HumanEval+59.1%#30 of 45, top 67%EvalPlus
MBPP+61.6%#23 of 38, top 61%EvalPlus

Reasoning

Qwen1.5-72B Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Hard Prompts1148#255 of 297, top 86%LMArena2026-10-08

Math

Qwen1.5-72B Math benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Math1164#245 of 285, top 86%LMArena2026-10-08

Knowledge

Qwen1.5-72B Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
GPQA Diamond28.8%#171 of 186, top 92%Epoch AI2025-01-27
LMArena Expert1136#243 of 273, top 90%LMArena2026-10-08

Multilingual

Qwen1.5-72B Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1135#253 of 297, top 86%LMArena2026-10-08
LMArena Chinese1186#237 of 285, top 84%LMArena2026-10-08
LMArena French1159#203 of 223, top 92%LMArena2026-10-08
LMArena German1084#213 of 231, top 93%LMArena2026-10-08
LMArena Japanese1061#189 of 211, top 90%LMArena2026-10-08
LMArena Korean1050#194 of 213, top 92%LMArena2026-10-08
LMArena Russian1104#257 of 283, top 91%LMArena2026-10-08
LMArena Spanish1110#215 of 226, top 96%LMArena2026-10-08

Instruction Following

Qwen1.5-72B Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Instruction Following1141#255 of 298, top 86%LMArena2026-10-08

Long Context

Qwen1.5-72B Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Longer Query1157#252 of 291, top 87%LMArena2026-10-08

Writing & Preference

Qwen1.5-72B Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1166#253 of 297, top 86%LMArena2026-10-08
LMArena Creative Writing1137#255 of 295, top 87%LMArena2026-10-08
LMArena Multi-Turn1160#248 of 295, top 85%LMArena2026-10-08

Compare Qwen1.5-72B

Other Alibaba (Qwen) models

Frequently asked questions

How good is Qwen1.5-72B?

Qwen1.5-72B by Alibaba (Qwen) ranks 285th of 354 ranked models on the Noometry Index as of October 2026, with a score of 30.8. Its strongest category is reasoning, where it ranks 203rd.

Is Qwen1.5-72B open source?

Yes. Qwen1.5-72B's weights are downloadable; check the license for commercial terms.

What are Qwen1.5-72B's strengths and weaknesses?

Relative to other ranked models, Qwen1.5-72B places best in reasoning, math, coding and lowest in knowledge, multilingual, instruction following.

What is Qwen1.5-72B best at?

Its best category is reasoning, where it ranks 203rd on Noometry.