Model comparison
Ministral 8B vs Qwen1.5 4b Chat
Ministral 8B and Qwen1.5 4b Chat score almost the same on the Noometry Index (28.2 vs 28.8), so choose on price, context window or the category you care about most.
Last verified . 12 shared benchmarks.
Summary
- They share 12 benchmarks with published results for both. Ministral 8B scores higher in 5 categories and Qwen1.5 4b Chat in 3 categories; 7 gaps are clear of the uncertainty.
- The widest gap is in writing & preference, where Ministral 8B leads 39.6 to 23.8.
Side by side
| Ministral 8B | Qwen1.5 4b Chat | |
|---|---|---|
| Provider | Mistral AI | Alibaba (Qwen) |
| Noometry Index | 28.2 | 28.8 |
| Released | 2024-10-01 | — |
| Weights | Open | Open |
| Context window | 262K | — |
| Max output | 262K | — |
| Input $ / M tokens | $0.15 | — |
| Output $ / M tokens | $0.15 | — |
| Results tracked | 17 | 13 |
Sponsored placements are available on pages like this one. Advertise on Noometry
Category by category
Coding Ministral 8B leads
Ministral 8B: 35.0 (#230), Qwen1.5 4b Chat: 29.1 (#308)
| Benchmark | Ministral 8B | Qwen1.5 4b Chat |
|---|---|---|
| LMArena Coding | 1202 | 999 |
Agentic & Tool Use Not comparable
Ministral 8B: 16.4 (#148), Qwen1.5 4b Chat: —
| Benchmark | Ministral 8B | Qwen1.5 4b Chat |
|---|---|---|
| Berkeley Function Calling Leaderboard | 11.1% | — |
Reasoning Too close to call
Ministral 8B: 18.4 (#281), Qwen1.5 4b Chat: 18.5 (#279)
| Benchmark | Ministral 8B | Qwen1.5 4b Chat |
|---|---|---|
| LMArena Hard Prompts | 1191 | 976 |
| DTBench | 45.7% | — |
Math Qwen1.5 4b Chat leads
Ministral 8B: 25.7 (#267), Qwen1.5 4b Chat: 30.4 (#234)
| Benchmark | Ministral 8B | Qwen1.5 4b Chat |
|---|---|---|
| LMArena Math | 1188 | 1026 |
| MATH Level 5 | 14.9% | — |
Knowledge Qwen1.5 4b Chat leads
Ministral 8B: 12.6 (#297), Qwen1.5 4b Chat: 26.7 (#255)
| Benchmark | Ministral 8B | Qwen1.5 4b Chat |
|---|---|---|
| LMArena Expert | 1170 | 980 |
| GPQA Diamond | 27.1% | — |
| Vectara Hallucination Rate | 7.4% | — |
Multilingual Ministral 8B leads
Ministral 8B: 35.1 (#247), Qwen1.5 4b Chat: 24.1 (#290)
| Benchmark | Ministral 8B | Qwen1.5 4b Chat |
|---|---|---|
| LMArena Non-English | 1165 | 979 |
| LMArena Chinese | 1193 | 1024 |
| LMArena Russian | 1195 | 952 |
| LMArena German | — | 902 |
Instruction Following Ministral 8B leads
Ministral 8B: 60.5 (#250), Qwen1.5 4b Chat: 49.0 (#300)
| Benchmark | Ministral 8B | Qwen1.5 4b Chat |
|---|---|---|
| LMArena Instruction Following | 1161 | 978 |
Long Context Ministral 8B leads
Ministral 8B: 36.7 (#227), Qwen1.5 4b Chat: 30.1 (#290)
| Benchmark | Ministral 8B | Qwen1.5 4b Chat |
|---|---|---|
| LMArena Longer Query | 1212 | 988 |
Writing & Preference Ministral 8B leads
Ministral 8B: 39.6 (#246), Qwen1.5 4b Chat: 23.8 (#309)
| Benchmark | Ministral 8B | Qwen1.5 4b Chat |
|---|---|---|
| LMArena Text | 1191 | 997 |
| LMArena Creative Writing | 1175 | 969 |
| LMArena Multi-Turn | 1166 | 977 |
Frequently asked questions
Is Ministral 8B better than Qwen1.5 4b Chat?
Ministral 8B and Qwen1.5 4b Chat score almost the same on the Noometry Index (28.2 vs 28.8), so choose on price, context window or the category you care about most.
Is Ministral 8B or Qwen1.5 4b Chat better for coding?
Ministral 8B scores higher on coding benchmarks: 35.0 versus 29.1 in the Noometry coding category.
How many benchmarks do Ministral 8B and Qwen1.5 4b Chat share?
12 benchmarks have published results for both models. Ministral 8B has 17 scored results on Noometry and Qwen1.5 4b Chat has 13.