Alibaba (Qwen), open weights
Qwen1.5 4b Chat
Qwen1.5 4b Chat by Alibaba (Qwen) ranks 322nd of 354 ranked models on the Noometry Index as of October 2026, with a score of 28.8. Its strongest category is math, where it ranks 234th.
Last verified
Specifications
- Noometry rank
- #322 of 354
- Index score
- 28.8
- Evidence
- Confirmed 13 results
- Provider
Alibaba (Qwen)
- Released
- Unknown
- Weights
- Open weights
- Reasoning
- Unknown
- Context window
- —
- Max output
- —
- Input price
- Not listed
- Output price
- Not listed
- Blended price
- Not listed
- Output speed
- Not measured
- Value
- Not ranked
- Knowledge cutoff
- Unknown
Category scores
Each category score combines every public result we have in that category.
- Coding 29.1
- Reasoning 18.5
- Math 30.4
- Knowledge 26.7
- Multilingual 24.1
- Instruction Following 49.0
- Long Context 30.1
- Writing & Preference 23.8
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 29.1 | #308 | 1 |
| Reasoning | 18.5 | #279 | 1 |
| Math | 30.4 | #234 | 1 |
| Knowledge | 26.7 | #255 | 1 |
| Multilingual | 24.1 | #290 | 1 |
| Instruction Following | 49.0 | #300 | 1 |
| Long Context | 30.1 | #290 | 1 |
| Writing & Preference | 23.8 | #309 | 3 |
Strengths and weaknesses
Categories where Qwen1.5 4b Chat places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Writing & Preference | 23.8 | −30.0 | #309 of 312, top 100% |
| Instruction Following | 49.0 | −22.3 | #300 of 305, top 99% |
| Long Context | 30.1 | −10.8 | #290 of 296, top 98% |
Closest competitors
The models ranked just above and below Qwen1.5 4b Chat. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Granite 4.0 Micro | #318 | 29.0 | $0.0408 | — | Compare |
| Claude 3 Sonnet | #319 | 29.0 | — | — | Compare |
| Qwen2.5 7B Instruct | #320 | 29.0 | $0.31 | — | Compare |
| Llama 3.2 3B | #321 | 28.9 | $0.12 | — | Compare |
| Llama 3-70B | #323 | 28.8 | — | 104 | Compare |
| GPT-4o | #324 | 28.6 | $4.38 | — | Compare |
| Ministral 8B | #325 | 28.2 | $0.15 | — | Compare |
| Gemma 3 4B | #326 | 28.1 | $0.05 | 72 | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Coding | 999 | #291 of 294, top 99% | LMArena | 2026-10-08 |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Hard Prompts | 976 | #295 of 297, top 100% | LMArena | 2026-10-08 |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Math | 1026 | #280 of 285, top 99% | LMArena | 2026-10-08 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Expert | 980 | #272 of 273, top 100% | LMArena | 2026-10-08 |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 979 | #290 of 297, top 98% | LMArena | 2026-10-08 | |
| LMArena Chinese | 1024 | #270 of 285, top 95% | LMArena | 2026-10-08 | |
| LMArena German | 902 | #231 of 231, top 100% | LMArena | 2026-10-08 | |
| LMArena Russian | 952 | #279 of 283, top 99% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Instruction Following | 978 | #294 of 298, top 99% | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 988 | #290 of 291, top 100% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 997 | #295 of 297, top 100% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 969 | #293 of 295, top 100% | LMArena | 2026-10-08 | |
| LMArena Multi-Turn | 977 | #290 of 295, top 99% | LMArena | 2026-10-08 |
Compare Qwen1.5 4b Chat
- Qwen1.5 4b Chat vs Llama 3.2 3B
- Qwen1.5 4b Chat vs Llama 3-70B
- Qwen1.5 4b Chat vs Qwen2.5 7B Instruct
- Qwen1.5 4b Chat vs GPT-4o
- Qwen1.5 4b Chat vs Claude 3 Sonnet
- Qwen1.5 4b Chat vs Ministral 8B
- Qwen1.5 4b Chat vs GPT-6 Astra
- Qwen1.5 4b Chat vs Claude Fable 5.1
- Qwen1.5 4b Chat vs Gemini 3.8 Flash
- Qwen1.5 4b Chat vs Kimi K3
- Qwen1.5 4b Chat vs Grok 4.6
- Qwen1.5 4b Chat vs GLM-5.3
- Qwen1.5 4b Chat vs Muse Spark 1.3
- Qwen1.5 4b Chat vs DeepSeek V4 Pro
Other Alibaba (Qwen) models
- Qwen3.8 Max56.8
- Qwen3.7 Max51.5
- Qwen3.6 Max Preview51.5
- Qwen3.6 Plus47.5
- Qwen3.5 397B-A17B46.0
- Qwen3.8 27B46.0
- Qwen3.5 Max Preview45.3
- Qwen3.7 Plus45.3
Frequently asked questions
How good is Qwen1.5 4b Chat?
Qwen1.5 4b Chat by Alibaba (Qwen) ranks 322nd of 354 ranked models on the Noometry Index as of October 2026, with a score of 28.8. Its strongest category is math, where it ranks 234th.
Is Qwen1.5 4b Chat open source?
Yes. Qwen1.5 4b Chat's weights are downloadable; check the license for commercial terms.
What are Qwen1.5 4b Chat's strengths and weaknesses?
Relative to other ranked models, Qwen1.5 4b Chat places best in math, reasoning, knowledge and lowest in writing & preference, instruction following, long context.
What is Qwen1.5 4b Chat best at?
Its best category is math, where it ranks 234th on Noometry.