Alibaba (Qwen), open weights
Qwen1.5-72B
Qwen1.5-72B by Alibaba (Qwen) ranks 285th of 354 ranked models on the Noometry Index as of October 2026, with a score of 30.8. Its strongest category is reasoning, where it ranks 203rd.
Last verified
Specifications
- Noometry rank
- #285 of 354
- Index score
- 30.8
- Evidence
- Confirmed 22 results
- Provider
Alibaba (Qwen)
- Released
- February 4, 2024
- Weights
- Open weights
- Reasoning
- Unknown
- Context window
- —
- Max output
- —
- Input price
- Not listed
- Output price
- Not listed
- Blended price
- Not listed
- Output speed
- Not measured
- Value
- Not ranked
- Knowledge cutoff
- Unknown
Category scores
Each category score combines every public result we have in that category.
- Coding 31.9
- Reasoning 22.2
- Math 33.2
- Knowledge 11.5
- Multilingual 33.2
- Instruction Following 59.3
- Long Context 35.1
- Writing & Preference 37.3
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 31.9 | #277 | 3 |
| Reasoning | 22.2 | #203 | 1 |
| Math | 33.2 | #205 | 1 |
| Knowledge | 11.5 | #300 | 2 |
| Multilingual | 33.2 | #253 | 1 |
| Instruction Following | 59.3 | #256 | 1 |
| Long Context | 35.1 | #243 | 1 |
| Writing & Preference | 37.3 | #258 | 3 |
Strengths and weaknesses
Categories where Qwen1.5-72B places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Knowledge | 11.5 | −25.8 | #300 of 314, top 96% |
| Multilingual | 33.2 | −14.2 | #253 of 297, top 86% |
| Instruction Following | 59.3 | −12.0 | #256 of 305, top 84% |
Closest competitors
The models ranked just above and below Qwen1.5-72B. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Amazon Nova Pro | #281 | 31.0 | $1.40 | — | Compare |
| Llama 4 Maverick | #282 | 30.9 | $0.30 | 456 | Compare |
| Phi-4 Mini | #283 | 30.9 | $0.13 | — | Compare |
| Gemma 3 27B | #284 | 30.8 | $0.10 | 62 | Compare |
| Granite 3.0 2b Instruct | #286 | 30.8 | — | — | Compare |
| Codellama 34b Instruct | #287 | 30.8 | — | — | Compare |
| Llama 3.1-405B | #288 | 30.7 | — | 78 | Compare |
| Yi-1.5-34B | #289 | 30.6 | — | — | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| BigCodeBench Instruct | 33.2% | #50 of 64, top 79% | BigCodeBench | 2024-04-26 | |
| LMArena Coding | 1165 | #253 of 294, top 87% | LMArena | 2026-10-08 | |
| BigCodeBench Complete | 40.3% | #53 of 66, top 81% | BigCodeBench | 2024-04-26 | |
| HumanEval+ | 59.1% | #30 of 45, top 67% | EvalPlus | ||
| MBPP+ | 61.6% | #23 of 38, top 61% | EvalPlus |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Hard Prompts | 1148 | #255 of 297, top 86% | LMArena | 2026-10-08 |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Math | 1164 | #245 of 285, top 86% | LMArena | 2026-10-08 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| GPQA Diamond | 28.8% | #171 of 186, top 92% | Epoch AI | 2025-01-27 | |
| LMArena Expert | 1136 | #243 of 273, top 90% | LMArena | 2026-10-08 |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1135 | #253 of 297, top 86% | LMArena | 2026-10-08 | |
| LMArena Chinese | 1186 | #237 of 285, top 84% | LMArena | 2026-10-08 | |
| LMArena French | 1159 | #203 of 223, top 92% | LMArena | 2026-10-08 | |
| LMArena German | 1084 | #213 of 231, top 93% | LMArena | 2026-10-08 | |
| LMArena Japanese | 1061 | #189 of 211, top 90% | LMArena | 2026-10-08 | |
| LMArena Korean | 1050 | #194 of 213, top 92% | LMArena | 2026-10-08 | |
| LMArena Russian | 1104 | #257 of 283, top 91% | LMArena | 2026-10-08 | |
| LMArena Spanish | 1110 | #215 of 226, top 96% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Instruction Following | 1141 | #255 of 298, top 86% | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 1157 | #252 of 291, top 87% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1166 | #253 of 297, top 86% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1137 | #255 of 295, top 87% | LMArena | 2026-10-08 | |
| LMArena Multi-Turn | 1160 | #248 of 295, top 85% | LMArena | 2026-10-08 |
Compare Qwen1.5-72B
- Qwen1.5-72B vs Gemma 3 27B
- Qwen1.5-72B vs Granite 3.0 2b Instruct
- Qwen1.5-72B vs Phi-4 Mini
- Qwen1.5-72B vs Codellama 34b Instruct
- Qwen1.5-72B vs Llama 4 Maverick
- Qwen1.5-72B vs Llama 3.1-405B
- Qwen1.5-72B vs GPT-6 Astra
- Qwen1.5-72B vs Claude Fable 5.1
- Qwen1.5-72B vs Gemini 3.8 Flash
- Qwen1.5-72B vs Kimi K3
- Qwen1.5-72B vs Grok 4.6
- Qwen1.5-72B vs GLM-5.3
- Qwen1.5-72B vs Muse Spark 1.3
- Qwen1.5-72B vs DeepSeek V4 Pro
Other Alibaba (Qwen) models
- Qwen3.8 Max56.8
- Qwen3.7 Max51.5
- Qwen3.6 Max Preview51.5
- Qwen3.6 Plus47.5
- Qwen3.5 397B-A17B46.0
- Qwen3.8 27B46.0
- Qwen3.5 Max Preview45.3
- Qwen3.7 Plus45.3
Frequently asked questions
How good is Qwen1.5-72B?
Qwen1.5-72B by Alibaba (Qwen) ranks 285th of 354 ranked models on the Noometry Index as of October 2026, with a score of 30.8. Its strongest category is reasoning, where it ranks 203rd.
Is Qwen1.5-72B open source?
Yes. Qwen1.5-72B's weights are downloadable; check the license for commercial terms.
What are Qwen1.5-72B's strengths and weaknesses?
Relative to other ranked models, Qwen1.5-72B places best in reasoning, math, coding and lowest in knowledge, multilingual, instruction following.
What is Qwen1.5-72B best at?
Its best category is reasoning, where it ranks 203rd on Noometry.