Alibaba (Qwen), open weights
Qwen1.5-14B
Qwen1.5-14B by Alibaba (Qwen) ranks 253rd of 354 ranked models on the Noometry Index as of October 2026, with a score of 32.7. Its strongest category is math, where it ranks 215th.
Last verified
Specifications
- Noometry rank
- #253 of 354
- Index score
- 32.7
- Evidence
- Confirmed 17 results
- Provider
Alibaba (Qwen)
- Released
- February 4, 2024
- Weights
- Open weights
- Reasoning
- Unknown
- Context window
- —
- Max output
- —
- Input price
- Not listed
- Output price
- Not listed
- Blended price
- Not listed
- Output speed
- Not measured
- Value
- Not ranked
- Knowledge cutoff
- Unknown
Category scores
Each category score combines every public result we have in that category.
- Coding 33.1
- Reasoning 21.4
- Math 32.4
- Knowledge 29.8
- Multilingual 30.7
- Instruction Following 56.8
- Long Context 33.7
- Writing & Preference 33.6
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 33.1 | #263 | 1 |
| Reasoning | 21.4 | #223 | 1 |
| Math | 32.4 | #215 | 1 |
| Knowledge | 29.8 | #232 | 1 |
| Multilingual | 30.7 | #262 | 1 |
| Instruction Following | 56.8 | #271 | 1 |
| Long Context | 33.7 | #257 | 1 |
| Writing & Preference | 33.6 | #276 | 3 |
Strengths and weaknesses
Categories where Qwen1.5-14B places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Instruction Following | 56.8 | −14.4 | #271 of 305, top 89% |
| Writing & Preference | 33.6 | −20.2 | #276 of 312, top 89% |
| Multilingual | 30.7 | −16.7 | #262 of 297, top 89% |
Closest competitors
The models ranked just above and below Qwen1.5-14B. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Wizardlm 70b | #249 | 33.0 | — | — | Compare |
| Phi 3 Medium 4k Instruct | #250 | 33.0 | — | — | Compare |
| Tulu 3 (Tülu 3) 70B | #251 | 33.0 | — | — | Compare |
| DeepSeek-R1-Distill-Qwen-14B | #252 | 32.7 | — | — | Compare |
| Olmo 2 0325 32b Instruct | #254 | 32.7 | — | — | Compare |
| gpt-oss-20b | #255 | 32.5 | $0.036 | 96 | Compare |
| Laguna M.1 | #256 | 32.5 | — | — | Compare |
| Command R+ | #257 | 32.4 | $4.38 | — | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Coding | 1138 | #259 of 294, top 89% | LMArena | 2026-10-08 |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Hard Prompts | 1113 | #262 of 297, top 89% | LMArena | 2026-10-08 |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Math | 1125 | #260 of 285, top 92% | LMArena | 2026-10-08 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Expert | 1094 | #251 of 273, top 92% | LMArena | 2026-10-08 | |
| MMLU | 68.6% | #51 of 81, top 63% | Epoch AI |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1095 | #262 of 297, top 89% | LMArena | 2026-10-08 | |
| LMArena Chinese | 1147 | #247 of 285, top 87% | LMArena | 2026-10-08 | |
| LMArena French | 1116 | #211 of 223, top 95% | LMArena | 2026-10-08 | |
| LMArena German | 1043 | #221 of 231, top 96% | LMArena | 2026-10-08 | |
| LMArena Japanese | 1019 | #197 of 211, top 94% | LMArena | 2026-10-08 | |
| LMArena Russian | 1046 | #269 of 283, top 96% | LMArena | 2026-10-08 | |
| LMArena Spanish | 1085 | #219 of 226, top 97% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Instruction Following | 1102 | #267 of 298, top 90% | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 1113 | #264 of 291, top 91% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1128 | #264 of 297, top 89% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1091 | #268 of 295, top 91% | LMArena | 2026-10-08 | |
| LMArena Multi-Turn | 1110 | #263 of 295, top 90% | LMArena | 2026-10-08 |
Compare Qwen1.5-14B
- Qwen1.5-14B vs DeepSeek-R1-Distill-Qwen-14B
- Qwen1.5-14B vs Olmo 2 0325 32b Instruct
- Qwen1.5-14B vs Tulu 3 (Tülu 3) 70B
- Qwen1.5-14B vs gpt-oss-20b
- Qwen1.5-14B vs Phi 3 Medium 4k Instruct
- Qwen1.5-14B vs Laguna M.1
- Qwen1.5-14B vs GPT-6 Astra
- Qwen1.5-14B vs Claude Fable 5.1
- Qwen1.5-14B vs Gemini 3.8 Flash
- Qwen1.5-14B vs Kimi K3
- Qwen1.5-14B vs Grok 4.6
- Qwen1.5-14B vs GLM-5.3
- Qwen1.5-14B vs Muse Spark 1.3
- Qwen1.5-14B vs DeepSeek V4 Pro
Other Alibaba (Qwen) models
- Qwen3.8 Max56.8
- Qwen3.7 Max51.5
- Qwen3.6 Max Preview51.5
- Qwen3.6 Plus47.5
- Qwen3.5 397B-A17B46.0
- Qwen3.8 27B46.0
- Qwen3.5 Max Preview45.3
- Qwen3.7 Plus45.3
Frequently asked questions
How good is Qwen1.5-14B?
Qwen1.5-14B by Alibaba (Qwen) ranks 253rd of 354 ranked models on the Noometry Index as of October 2026, with a score of 32.7. Its strongest category is math, where it ranks 215th.
Is Qwen1.5-14B open source?
Yes. Qwen1.5-14B's weights are downloadable; check the license for commercial terms.
What are Qwen1.5-14B's strengths and weaknesses?
Relative to other ranked models, Qwen1.5-14B places best in reasoning, math, knowledge and lowest in instruction following, writing & preference, multilingual.
What is Qwen1.5-14B best at?
Its best category is math, where it ranks 215th on Noometry.