Model comparison

Gemma 3 27B vs Granite 3.0 2b Instruct

Gemma 3 27B and Granite 3.0 2b Instruct score almost the same on the Noometry Index (30.8 vs 30.8), so choose on price, context window or the category you care about most.

Last verified . 12 shared benchmarks.

Gemma 3 27B Google

30.8

Rank #284 Confirmed

Granite 3.0 2b Instruct IBM

30.8

Rank #286 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Gemma 3 27B scores higher in 3 categories and Granite 3.0 2b Instruct in 5 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Gemma 3 27B leads 52.5 to 29.6.

Side by side

Gemma 3 27B and Granite 3.0 2b Instruct specifications
Gemma 3 27BGranite 3.0 2b Instruct
ProviderGoogleIBM
Noometry Index30.830.8
Released2025-03-11—
WeightsOpenOpen
Context window131K—
Max output8K—
Input $ / M tokens$0.08—
Output $ / M tokens$0.16—
Results tracked4313

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Granite 3.0 2b Instruct leads

Gemma 3 27B: 22.5 (#334), Granite 3.0 2b Instruct: 28.3 (#316)

Coding benchmarks
BenchmarkGemma 3 27BGranite 3.0 2b Instruct
LMArena Coding13221090
Aider Polyglot4.9%—
SciCode21.2%—
BigCodeBench Instruct—20.5%
LiveBench Coding39.9%—

Agentic & Tool Use Not comparable

Gemma 3 27B: 25.1 (#110), Granite 3.0 2b Instruct: —

Agentic & Tool Use benchmarks
BenchmarkGemma 3 27BGranite 3.0 2b Instruct
Berkeley Function Calling Leaderboard29.5%—

Reasoning Granite 3.0 2b Instruct leads

Gemma 3 27B: 16.7 (#301), Granite 3.0 2b Instruct: 20.5 (#235)

Reasoning benchmarks
BenchmarkGemma 3 27BGranite 3.0 2b Instruct
LMArena Hard Prompts13401073
Kagi LLM Benchmark40.4%—
CritPt0%—
Chess Puzzles0%—
LiveBench Reasoning43.8%—
DTBench52.5%—
LiveBench Data Analysis51.5%—
LMCA12.3%—
Epoch Capabilities Index130.04—
LiveBench50%—

Math Granite 3.0 2b Instruct leads

Gemma 3 27B: 25.9 (#265), Granite 3.0 2b Instruct: 32.2 (#217)

Math benchmarks
BenchmarkGemma 3 27BGranite 3.0 2b Instruct
LMArena Math13121117
OTIS Mock AIME 2024-202522.5%—
LiveBench Math55.4%—
MATH Level 574%—

Knowledge Granite 3.0 2b Instruct leads

Gemma 3 27B: 25.5 (#261), Granite 3.0 2b Instruct: 29.0 (#241)

Knowledge benchmarks
BenchmarkGemma 3 27BGranite 3.0 2b Instruct
LMArena Expert13041064
GPQA Diamond47.7%—
Confabulations40.3%—
Vectara Hallucination Rate7.4%—

Multimodal Not comparable

Gemma 3 27B: 32.6 (#100), Granite 3.0 2b Instruct: —

Multimodal benchmarks
BenchmarkGemma 3 27BGranite 3.0 2b Instruct
LMArena Vision1164—
GeoBench52%—

Multilingual Gemma 3 27B leads

Gemma 3 27B: 46.9 (#155), Granite 3.0 2b Instruct: 27.0 (#278)

Multilingual benchmarks
BenchmarkGemma 3 27BGranite 3.0 2b Instruct
LMArena Non-English13341033
LMArena Chinese13461070
LMArena Russian13491045
LMArena French1368—
LMArena German1362—
LMArena Japanese1287—
LMArena Korean1308—
LMArena Spanish1349—

Instruction Following Gemma 3 27B leads

Gemma 3 27B: 70.6 (#160), Granite 3.0 2b Instruct: 53.9 (#284)

Instruction Following benchmarks
BenchmarkGemma 3 27BGranite 3.0 2b Instruct
LMArena Instruction Following13211056
LiveBench Instruction Following74.9%—

Long Context Granite 3.0 2b Instruct leads

Gemma 3 27B: 27.6 (#293), Granite 3.0 2b Instruct: 32.5 (#268)

Long Context benchmarks
BenchmarkGemma 3 27BGranite 3.0 2b Instruct
LMArena Longer Query13331070
Fiction.LiveBench33.3%—

Writing & Preference Gemma 3 27B leads

Gemma 3 27B: 52.5 (#168), Granite 3.0 2b Instruct: 29.6 (#292)

Writing & Preference benchmarks
BenchmarkGemma 3 27BGranite 3.0 2b Instruct
LMArena Text13581080
LMArena Creative Writing13461046
LMArena Multi-Turn13451053
Short-Story Creative Writing79.9%—
EQ-Bench Creative Writing1266—
LiveBench Language34.6%—

Frequently asked questions

Is Gemma 3 27B better than Granite 3.0 2b Instruct?

Gemma 3 27B and Granite 3.0 2b Instruct score almost the same on the Noometry Index (30.8 vs 30.8), so choose on price, context window or the category you care about most.

Is Gemma 3 27B or Granite 3.0 2b Instruct better for coding?

Granite 3.0 2b Instruct scores higher on coding benchmarks: 28.3 versus 22.5 in the Noometry coding category.

How many benchmarks do Gemma 3 27B and Granite 3.0 2b Instruct share?

12 benchmarks have published results for both models. Gemma 3 27B has 43 scored results on Noometry and Granite 3.0 2b Instruct has 13.

Related comparisons

Go deeper