Model comparison

Claude 3 Sonnet vs Granite 4.0 Micro

Claude 3 Sonnet and Granite 4.0 Micro score almost the same on the Noometry Index (29.0 vs 29.0), so choose on price, context window or the category you care about most.

Last verified . 2 shared benchmarks.

Claude 3 Sonnet Anthropic

29.0

Rank #319 Confirmed

Granite 4.0 Micro IBM

29.0

Rank #318 Confirmed

Summary

  • They share 2 benchmarks with published results for both. Claude 3 Sonnet scores higher in 2 categories and Granite 4.0 Micro in 3 categories; 5 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Claude 3 Sonnet leads 21.1 to 9.9.
  • The biggest single-benchmark swing is GPQA Diamond: 40.6% for Claude 3 Sonnet and 28.3% for Granite 4.0 Micro.
  • Granite 4.0 Micro has downloadable open weights; the other is API-only.

Side by side

Claude 3 Sonnet and Granite 4.0 Micro specifications
Claude 3 SonnetGranite 4.0 Micro
ProviderAnthropicIBM
Noometry Index29.029.0
Released2024-02-292025-10-02
WeightsProprietaryOpen
Context window—131K
Max output—118K
Input $ / M tokens—$0.017
Output $ / M tokens—$0.11
Results tracked308

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Claude 3 Sonnet: 29.6 (#302), Granite 4.0 Micro: —

Coding benchmarks
BenchmarkClaude 3 SonnetGranite 4.0 Micro
WeirdML10.2%—
BigCodeBench Instruct42.7%—
LMArena Coding1223—
BigCodeBench Complete53.8%—
HumanEval+64%—
MBPP+69.3%—

Reasoning Claude 3 Sonnet leads

Claude 3 Sonnet: 20.5 (#237), Granite 4.0 Micro: 19.2 (#265)

Reasoning benchmarks
BenchmarkClaude 3 SonnetGranite 4.0 Micro
Chess Puzzles—0%
LMArena Hard Prompts1197—
DTBench53.6%—
Epoch Capabilities Index120.7—
WinoGrande75.1%—

Math Granite 4.0 Micro leads

Claude 3 Sonnet: 10.7 (#310), Granite 4.0 Micro: 12.0 (#307)

Math benchmarks
BenchmarkClaude 3 SonnetGranite 4.0 Micro
OTIS Mock AIME 2024-20252.5%2.8%
Omni-MATH—20.9%
LMArena Math1213—
MATH Level 518.2%—

Knowledge Claude 3 Sonnet leads

Claude 3 Sonnet: 21.1 (#276), Granite 4.0 Micro: 9.9 (#304)

Knowledge benchmarks
BenchmarkClaude 3 SonnetGranite 4.0 Micro
GPQA Diamond40.6%28.3%
MMLU-Pro—39.5%
GPQA (HELM)—30.7%
LMArena Expert1173—
MMLU75.9%—

Multimodal Not comparable

Claude 3 Sonnet: 25.2 (#125), Granite 4.0 Micro: —

Multimodal benchmarks
BenchmarkClaude 3 SonnetGranite 4.0 Micro
LMArena Vision984—

Multilingual Not comparable

Claude 3 Sonnet: 37.8 (#234), Granite 4.0 Micro: —

Multilingual benchmarks
BenchmarkClaude 3 SonnetGranite 4.0 Micro
LMArena Non-English1205—
LMArena Chinese1189—
LMArena French1229—
LMArena German1204—
LMArena Japanese1131—
LMArena Korean1128—
LMArena Russian1227—
LMArena Spanish1204—

Instruction Following Granite 4.0 Micro leads

Claude 3 Sonnet: 62.8 (#235), Granite 4.0 Micro: 69.9 (#169)

Instruction Following benchmarks
BenchmarkClaude 3 SonnetGranite 4.0 Micro
IFEval—84.9%
LMArena Instruction Following1199—

Long Context Not comparable

Claude 3 Sonnet: 36.7 (#228), Granite 4.0 Micro: —

Long Context benchmarks
BenchmarkClaude 3 SonnetGranite 4.0 Micro
LMArena Longer Query1211—

Writing & Preference Granite 4.0 Micro leads

Claude 3 Sonnet: 42.1 (#238), Granite 4.0 Micro: 46.7 (#216)

Writing & Preference benchmarks
BenchmarkClaude 3 SonnetGranite 4.0 Micro
LMArena Text1218—
LMArena Creative Writing1186—
WildBench—67%
LMArena Multi-Turn1227—

Frequently asked questions

Is Claude 3 Sonnet better than Granite 4.0 Micro?

Claude 3 Sonnet and Granite 4.0 Micro score almost the same on the Noometry Index (29.0 vs 29.0), so choose on price, context window or the category you care about most.

How many benchmarks do Claude 3 Sonnet and Granite 4.0 Micro share?

2 benchmarks have published results for both models. Claude 3 Sonnet has 30 scored results on Noometry and Granite 4.0 Micro has 8.

Related comparisons

Go deeper