Model comparison

Gemma 3n E4b IT vs Mercury

Gemma 3n E4b IT and Mercury score almost the same on the Noometry Index (37.3 vs 37.6), so choose on price, context window or the category you care about most.

Last verified . 9 shared benchmarks.

Gemma 3n E4b IT Google

37.3

Rank #206 Confirmed

Mercury Inception

37.6

Rank #199 Confirmed

Summary

  • They share 9 benchmarks with published results for both. Gemma 3n E4b IT scores higher in 5 categories and Mercury in 1 category; 4 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Gemma 3n E4b IT leads 50.1 to 46.2.
  • The biggest single-benchmark swing is Kagi LLM Benchmark: 31.5% for Gemma 3n E4b IT and 21.6% for Mercury.
  • Gemma 3n E4b IT has downloadable open weights; the other is API-only.

Side by side

Gemma 3n E4b IT and Mercury specifications
Gemma 3n E4b ITMercury
ProviderGoogleInception
Noometry Index37.337.6
Released——
WeightsOpenProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked189

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mercury leads

Gemma 3n E4b IT: 37.0 (#198), Mercury: 38.7 (#170)

Coding benchmarks
BenchmarkGemma 3n E4b ITMercury
LMArena Coding12681322

Reasoning Gemma 3n E4b IT leads

Gemma 3n E4b IT: 19.9 (#247), Mercury: 17.5 (#293)

Reasoning benchmarks
BenchmarkGemma 3n E4b ITMercury
Kagi LLM Benchmark31.5%21.6%
LMArena Hard Prompts12841285

Math Not comparable

Gemma 3n E4b IT: 35.1 (#188), Mercury: —

Math benchmarks
BenchmarkGemma 3n E4b ITMercury
LMArena Math1251—

Knowledge Not comparable

Gemma 3n E4b IT: 34.2 (#198), Mercury: —

Knowledge benchmarks
BenchmarkGemma 3n E4b ITMercury
LMArena Expert1246—

Multilingual Gemma 3n E4b IT leads

Gemma 3n E4b IT: 43.4 (#183), Mercury: 41.6 (#206)

Multilingual benchmarks
BenchmarkGemma 3n E4b ITMercury
LMArena Non-English12851260
LMArena Chinese1309—
LMArena French1330—
LMArena German1311—
LMArena Japanese1272—
LMArena Korean1259—
LMArena Russian1288—
LMArena Spanish1305—

Instruction Following Too close to call

Gemma 3n E4b IT: 66.1 (#210), Mercury: 65.2 (#224)

Instruction Following benchmarks
BenchmarkGemma 3n E4b ITMercury
LMArena Instruction Following12551239

Long Context Too close to call

Gemma 3n E4b IT: 38.7 (#191), Mercury: 38.4 (#198)

Long Context benchmarks
BenchmarkGemma 3n E4b ITMercury
LMArena Longer Query12761266

Writing & Preference Gemma 3n E4b IT leads

Gemma 3n E4b IT: 50.1 (#186), Mercury: 46.2 (#221)

Writing & Preference benchmarks
BenchmarkGemma 3n E4b ITMercury
LMArena Text13061282
LMArena Creative Writing12871191
LMArena Multi-Turn12761282

Frequently asked questions

Is Gemma 3n E4b IT better than Mercury?

Gemma 3n E4b IT and Mercury score almost the same on the Noometry Index (37.3 vs 37.6), so choose on price, context window or the category you care about most.

Is Gemma 3n E4b IT or Mercury better for coding?

Mercury scores higher on coding benchmarks: 38.7 versus 37.0 in the Noometry coding category.

How many benchmarks do Gemma 3n E4b IT and Mercury share?

9 benchmarks have published results for both models. Gemma 3n E4b IT has 18 scored results on Noometry and Mercury has 9.

Related comparisons

Go deeper