Model comparison

Gemma 3n E4b IT vs Grok 2 Mini 2024 08 13

Gemma 3n E4b IT and Grok 2 Mini 2024 08 13 score almost the same on the Noometry Index (37.3 vs 37.7), so choose on price, context window or the category you care about most.

Last verified . 17 shared benchmarks.

Gemma 3n E4b IT Google

37.3

Rank #206 Confirmed

Grok 2 Mini 2024 08 13 xAI

37.7

Rank #198 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Gemma 3n E4b IT scores higher in 5 categories and Grok 2 Mini 2024 08 13 in 3 categories; 3 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Grok 2 Mini 2024 08 13 leads 24.8 to 19.9.
  • Gemma 3n E4b IT has downloadable open weights; the other is API-only.

Side by side

Gemma 3n E4b IT and Grok 2 Mini 2024 08 13 specifications
Gemma 3n E4b ITGrok 2 Mini 2024 08 13
ProviderGooglexAI
Noometry Index37.337.7
Released—2024-08-13
WeightsOpenProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked1817

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Gemma 3n E4b IT: 37.0 (#198), Grok 2 Mini 2024 08 13: 37.0 (#199)

Coding benchmarks
BenchmarkGemma 3n E4b ITGrok 2 Mini 2024 08 13
LMArena Coding12681269

Reasoning Grok 2 Mini 2024 08 13 leads

Gemma 3n E4b IT: 19.9 (#247), Grok 2 Mini 2024 08 13: 24.8 (#159)

Reasoning benchmarks
BenchmarkGemma 3n E4b ITGrok 2 Mini 2024 08 13
LMArena Hard Prompts12841255
Kagi LLM Benchmark31.5%—

Math Too close to call

Gemma 3n E4b IT: 35.1 (#188), Grok 2 Mini 2024 08 13: 35.4 (#185)

Math benchmarks
BenchmarkGemma 3n E4b ITGrok 2 Mini 2024 08 13
LMArena Math12511265

Knowledge Too close to call

Gemma 3n E4b IT: 34.2 (#198), Grok 2 Mini 2024 08 13: 33.9 (#200)

Knowledge benchmarks
BenchmarkGemma 3n E4b ITGrok 2 Mini 2024 08 13
LMArena Expert12461238

Multilingual Gemma 3n E4b IT leads

Gemma 3n E4b IT: 43.4 (#183), Grok 2 Mini 2024 08 13: 41.4 (#208)

Multilingual benchmarks
BenchmarkGemma 3n E4b ITGrok 2 Mini 2024 08 13
LMArena Non-English12851257
LMArena Chinese13091262
LMArena French13301286
LMArena German13111273
LMArena Japanese12721213
LMArena Korean12591195
LMArena Russian12881261
LMArena Spanish13051279

Instruction Following Too close to call

Gemma 3n E4b IT: 66.1 (#210), Grok 2 Mini 2024 08 13: 65.5 (#220)

Instruction Following benchmarks
BenchmarkGemma 3n E4b ITGrok 2 Mini 2024 08 13
LMArena Instruction Following12551245

Long Context Too close to call

Gemma 3n E4b IT: 38.7 (#191), Grok 2 Mini 2024 08 13: 38.4 (#196)

Long Context benchmarks
BenchmarkGemma 3n E4b ITGrok 2 Mini 2024 08 13
LMArena Longer Query12761266

Writing & Preference Gemma 3n E4b IT leads

Gemma 3n E4b IT: 50.1 (#186), Grok 2 Mini 2024 08 13: 47.3 (#211)

Writing & Preference benchmarks
BenchmarkGemma 3n E4b ITGrok 2 Mini 2024 08 13
LMArena Text13061281
LMArena Creative Writing12871242
LMArena Multi-Turn12761265

Frequently asked questions

Is Gemma 3n E4b IT better than Grok 2 Mini 2024 08 13?

Gemma 3n E4b IT and Grok 2 Mini 2024 08 13 score almost the same on the Noometry Index (37.3 vs 37.7), so choose on price, context window or the category you care about most.

Is Gemma 3n E4b IT or Grok 2 Mini 2024 08 13 better for coding?

They score almost the same on coding (37.0 vs 37.0); test both on your own repository before choosing.

How many benchmarks do Gemma 3n E4b IT and Grok 2 Mini 2024 08 13 share?

17 benchmarks have published results for both models. Gemma 3n E4b IT has 18 scored results on Noometry and Grok 2 Mini 2024 08 13 has 17.

Related comparisons

Go deeper