Model comparison
Granite 4.0 Micro vs Qwen1.5 4b Chat
Granite 4.0 Micro and Qwen1.5 4b Chat score almost the same on the Noometry Index (29.0 vs 28.8), so choose on price, context window or the category you care about most.
Last verified . 0 shared benchmarks.
Summary
- The widest gap is in writing & preference, where Granite 4.0 Micro leads 46.7 to 23.8.
Side by side
| Granite 4.0 Micro | Qwen1.5 4b Chat | |
|---|---|---|
| Provider | IBM | Alibaba (Qwen) |
| Noometry Index | 29.0 | 28.8 |
| Released | 2025-10-02 | — |
| Weights | Open | Open |
| Context window | 131K | — |
| Max output | 118K | — |
| Input $ / M tokens | $0.017 | — |
| Output $ / M tokens | $0.11 | — |
| Results tracked | 8 | 13 |
Sponsored placements are available on pages like this one. Advertise on Noometry
Category by category
Coding Not comparable
Granite 4.0 Micro: —, Qwen1.5 4b Chat: 29.1 (#308)
| Benchmark | Granite 4.0 Micro | Qwen1.5 4b Chat |
|---|---|---|
| LMArena Coding | — | 999 |
Reasoning Too close to call
Granite 4.0 Micro: 19.2 (#265), Qwen1.5 4b Chat: 18.5 (#279)
| Benchmark | Granite 4.0 Micro | Qwen1.5 4b Chat |
|---|---|---|
| Chess Puzzles | 0% | — |
| LMArena Hard Prompts | — | 976 |
Math Qwen1.5 4b Chat leads
Granite 4.0 Micro: 12.0 (#307), Qwen1.5 4b Chat: 30.4 (#234)
| Benchmark | Granite 4.0 Micro | Qwen1.5 4b Chat |
|---|---|---|
| OTIS Mock AIME 2024-2025 | 2.8% | — |
| Omni-MATH | 20.9% | — |
| LMArena Math | — | 1026 |
Knowledge Qwen1.5 4b Chat leads
Granite 4.0 Micro: 9.9 (#304), Qwen1.5 4b Chat: 26.7 (#255)
| Benchmark | Granite 4.0 Micro | Qwen1.5 4b Chat |
|---|---|---|
| GPQA Diamond | 28.3% | — |
| MMLU-Pro | 39.5% | — |
| GPQA (HELM) | 30.7% | — |
| LMArena Expert | — | 980 |
Multilingual Not comparable
Granite 4.0 Micro: —, Qwen1.5 4b Chat: 24.1 (#290)
| Benchmark | Granite 4.0 Micro | Qwen1.5 4b Chat |
|---|---|---|
| LMArena Non-English | — | 979 |
| LMArena Chinese | — | 1024 |
| LMArena German | — | 902 |
| LMArena Russian | — | 952 |
Instruction Following Granite 4.0 Micro leads
Granite 4.0 Micro: 69.9 (#169), Qwen1.5 4b Chat: 49.0 (#300)
| Benchmark | Granite 4.0 Micro | Qwen1.5 4b Chat |
|---|---|---|
| IFEval | 84.9% | — |
| LMArena Instruction Following | — | 978 |
Long Context Not comparable
Granite 4.0 Micro: —, Qwen1.5 4b Chat: 30.1 (#290)
| Benchmark | Granite 4.0 Micro | Qwen1.5 4b Chat |
|---|---|---|
| LMArena Longer Query | — | 988 |
Writing & Preference Granite 4.0 Micro leads
Granite 4.0 Micro: 46.7 (#216), Qwen1.5 4b Chat: 23.8 (#309)
| Benchmark | Granite 4.0 Micro | Qwen1.5 4b Chat |
|---|---|---|
| LMArena Text | — | 997 |
| LMArena Creative Writing | — | 969 |
| WildBench | 67% | — |
| LMArena Multi-Turn | — | 977 |
Frequently asked questions
Is Granite 4.0 Micro better than Qwen1.5 4b Chat?
Granite 4.0 Micro and Qwen1.5 4b Chat score almost the same on the Noometry Index (29.0 vs 28.8), so choose on price, context window or the category you care about most.
How many benchmarks do Granite 4.0 Micro and Qwen1.5 4b Chat share?
0 benchmarks have published results for both models. Granite 4.0 Micro has 8 scored results on Noometry and Qwen1.5 4b Chat has 13.