Model comparison

Granite 4.0 Micro vs Phi 3 Mini 128k Instruct

Granite 4.0 Micro and Phi 3 Mini 128k Instruct score almost the same on the Noometry Index (29.0 vs 29.7), so choose on price, context window or the category you care about most.

Last verified . 0 shared benchmarks.

Granite 4.0 Micro IBM

29.0

Rank #318 Confirmed

Phi 3 Mini 128k Instruct Microsoft

29.7

Rank #305 Confirmed

Summary

  • The widest gap is in math, where Phi 3 Mini 128k Instruct leads 31.6 to 12.0.

Side by side

Granite 4.0 Micro and Phi 3 Mini 128k Instruct specifications
Granite 4.0 MicroPhi 3 Mini 128k Instruct
ProviderIBMMicrosoft
Noometry Index29.029.7
Released2025-10-022024-04-23
WeightsOpenOpen
Context window131K—
Max output118K—
Input $ / M tokens$0.017—
Output $ / M tokens$0.11—
Results tracked819

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Granite 4.0 Micro: —, Phi 3 Mini 128k Instruct: 28.8 (#312)

Coding benchmarks
BenchmarkGranite 4.0 MicroPhi 3 Mini 128k Instruct
BigCodeBench Instruct—29.6%
LMArena Coding—1039
BigCodeBench Complete—40.6%

Reasoning Too close to call

Granite 4.0 Micro: 19.2 (#265), Phi 3 Mini 128k Instruct: 19.6 (#256)

Reasoning benchmarks
BenchmarkGranite 4.0 MicroPhi 3 Mini 128k Instruct
Chess Puzzles0%—
LMArena Hard Prompts—1028

Math Phi 3 Mini 128k Instruct leads

Granite 4.0 Micro: 12.0 (#307), Phi 3 Mini 128k Instruct: 31.6 (#222)

Math benchmarks
BenchmarkGranite 4.0 MicroPhi 3 Mini 128k Instruct
OTIS Mock AIME 2024-20252.8%—
Omni-MATH20.9%—
LMArena Math—1089

Knowledge Phi 3 Mini 128k Instruct leads

Granite 4.0 Micro: 9.9 (#304), Phi 3 Mini 128k Instruct: 26.8 (#254)

Knowledge benchmarks
BenchmarkGranite 4.0 MicroPhi 3 Mini 128k Instruct
GPQA Diamond28.3%—
MMLU-Pro39.5%—
GPQA (HELM)30.7%—
LMArena Expert—984

Multilingual Not comparable

Granite 4.0 Micro: —, Phi 3 Mini 128k Instruct: 25.2 (#285)

Multilingual benchmarks
BenchmarkGranite 4.0 MicroPhi 3 Mini 128k Instruct
LMArena Non-English—1000
LMArena Chinese—1016
LMArena French—1039
LMArena German—1006
LMArena Japanese—899
LMArena Korean—856
LMArena Russian—1004
LMArena Spanish—1059

Instruction Following Granite 4.0 Micro leads

Granite 4.0 Micro: 69.9 (#169), Phi 3 Mini 128k Instruct: 51.8 (#294)

Instruction Following benchmarks
BenchmarkGranite 4.0 MicroPhi 3 Mini 128k Instruct
IFEval84.9%—
LMArena Instruction Following—1023

Long Context Not comparable

Granite 4.0 Micro: —, Phi 3 Mini 128k Instruct: 30.4 (#289)

Long Context benchmarks
BenchmarkGranite 4.0 MicroPhi 3 Mini 128k Instruct
LMArena Longer Query—996

Writing & Preference Granite 4.0 Micro leads

Granite 4.0 Micro: 46.7 (#216), Phi 3 Mini 128k Instruct: 27.1 (#301)

Writing & Preference benchmarks
BenchmarkGranite 4.0 MicroPhi 3 Mini 128k Instruct
LMArena Text—1050
LMArena Creative Writing—1024
WildBench67%—
LMArena Multi-Turn—989

Frequently asked questions

Is Granite 4.0 Micro better than Phi 3 Mini 128k Instruct?

Granite 4.0 Micro and Phi 3 Mini 128k Instruct score almost the same on the Noometry Index (29.0 vs 29.7), so choose on price, context window or the category you care about most.

How many benchmarks do Granite 4.0 Micro and Phi 3 Mini 128k Instruct share?

0 benchmarks have published results for both models. Granite 4.0 Micro has 8 scored results on Noometry and Phi 3 Mini 128k Instruct has 19.

Related comparisons

Go deeper