Model comparison

Granite 4.0 Micro vs Qwen1.5 4b Chat

Granite 4.0 Micro and Qwen1.5 4b Chat score almost the same on the Noometry Index (29.0 vs 28.8), so choose on price, context window or the category you care about most.

Last verified . 0 shared benchmarks.

Granite 4.0 Micro IBM

29.0

Rank #318 Confirmed

Qwen1.5 4b Chat Alibaba (Qwen)

28.8

Rank #322 Confirmed

Summary

  • The widest gap is in writing & preference, where Granite 4.0 Micro leads 46.7 to 23.8.

Side by side

Granite 4.0 Micro and Qwen1.5 4b Chat specifications
Granite 4.0 MicroQwen1.5 4b Chat
ProviderIBMAlibaba (Qwen)
Noometry Index29.028.8
Released2025-10-02—
WeightsOpenOpen
Context window131K—
Max output118K—
Input $ / M tokens$0.017—
Output $ / M tokens$0.11—
Results tracked813

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Granite 4.0 Micro: —, Qwen1.5 4b Chat: 29.1 (#308)

Coding benchmarks
BenchmarkGranite 4.0 MicroQwen1.5 4b Chat
LMArena Coding—999

Reasoning Too close to call

Granite 4.0 Micro: 19.2 (#265), Qwen1.5 4b Chat: 18.5 (#279)

Reasoning benchmarks
BenchmarkGranite 4.0 MicroQwen1.5 4b Chat
Chess Puzzles0%—
LMArena Hard Prompts—976

Math Qwen1.5 4b Chat leads

Granite 4.0 Micro: 12.0 (#307), Qwen1.5 4b Chat: 30.4 (#234)

Math benchmarks
BenchmarkGranite 4.0 MicroQwen1.5 4b Chat
OTIS Mock AIME 2024-20252.8%—
Omni-MATH20.9%—
LMArena Math—1026

Knowledge Qwen1.5 4b Chat leads

Granite 4.0 Micro: 9.9 (#304), Qwen1.5 4b Chat: 26.7 (#255)

Knowledge benchmarks
BenchmarkGranite 4.0 MicroQwen1.5 4b Chat
GPQA Diamond28.3%—
MMLU-Pro39.5%—
GPQA (HELM)30.7%—
LMArena Expert—980

Multilingual Not comparable

Granite 4.0 Micro: —, Qwen1.5 4b Chat: 24.1 (#290)

Multilingual benchmarks
BenchmarkGranite 4.0 MicroQwen1.5 4b Chat
LMArena Non-English—979
LMArena Chinese—1024
LMArena German—902
LMArena Russian—952

Instruction Following Granite 4.0 Micro leads

Granite 4.0 Micro: 69.9 (#169), Qwen1.5 4b Chat: 49.0 (#300)

Instruction Following benchmarks
BenchmarkGranite 4.0 MicroQwen1.5 4b Chat
IFEval84.9%—
LMArena Instruction Following—978

Long Context Not comparable

Granite 4.0 Micro: —, Qwen1.5 4b Chat: 30.1 (#290)

Long Context benchmarks
BenchmarkGranite 4.0 MicroQwen1.5 4b Chat
LMArena Longer Query—988

Writing & Preference Granite 4.0 Micro leads

Granite 4.0 Micro: 46.7 (#216), Qwen1.5 4b Chat: 23.8 (#309)

Writing & Preference benchmarks
BenchmarkGranite 4.0 MicroQwen1.5 4b Chat
LMArena Text—997
LMArena Creative Writing—969
WildBench67%—
LMArena Multi-Turn—977

Frequently asked questions

Is Granite 4.0 Micro better than Qwen1.5 4b Chat?

Granite 4.0 Micro and Qwen1.5 4b Chat score almost the same on the Noometry Index (29.0 vs 28.8), so choose on price, context window or the category you care about most.

How many benchmarks do Granite 4.0 Micro and Qwen1.5 4b Chat share?

0 benchmarks have published results for both models. Granite 4.0 Micro has 8 scored results on Noometry and Qwen1.5 4b Chat has 13.

Related comparisons

Go deeper