Model comparison

Gemma 1.1 7b IT vs Mistral Small 3

Gemma 1.1 7b IT and Mistral Small 3 score almost the same on the Noometry Index (31.3 vs 31.2), so choose on price, context window or the category you care about most.

Last verified . 16 shared benchmarks.

Gemma 1.1 7b IT Google

31.3

Rank #277 Confirmed

Mistral Small 3 Mistral AI

31.2

Rank #278 Confirmed

Summary

  • They share 16 benchmarks with published results for both. Gemma 1.1 7b IT scores higher in 3 categories and Mistral Small 3 in 5 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in math, where Gemma 1.1 7b IT leads 32.0 to 16.3.

Side by side

Gemma 1.1 7b IT and Mistral Small 3 specifications
Gemma 1.1 7b ITMistral Small 3
ProviderGoogleMistral AI
Noometry Index31.331.2
Released—2025-01-30
WeightsOpenOpen
Context window—33K
Max output—16K
Input $ / M tokens—$0.05
Output $ / M tokens—$0.08
Results tracked1924

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mistral Small 3 leads

Gemma 1.1 7b IT: 31.5 (#284), Mistral Small 3: 36.5 (#207)

Coding benchmarks
BenchmarkGemma 1.1 7b ITMistral Small 3
LMArena Coding10841246
BigCodeBench Instruct—45.3%
BigCodeBench Complete—50.4%
HumanEval+35.4%—
MBPP+45%—

Reasoning Gemma 1.1 7b IT leads

Gemma 1.1 7b IT: 20.5 (#238), Mistral Small 3: 18.9 (#273)

Reasoning benchmarks
BenchmarkGemma 1.1 7b ITMistral Small 3
LMArena Hard Prompts10711233
Chess Puzzles—0%
Epoch Capabilities Index—127.07

Math Gemma 1.1 7b IT leads

Gemma 1.1 7b IT: 32.0 (#220), Mistral Small 3: 16.3 (#295)

Math benchmarks
BenchmarkGemma 1.1 7b ITMistral Small 3
LMArena Math11071240
OTIS Mock AIME 2024-2025—6.7%

Knowledge Gemma 1.1 7b IT leads

Gemma 1.1 7b IT: 28.3 (#247), Mistral Small 3: 25.1 (#263)

Knowledge benchmarks
BenchmarkGemma 1.1 7b ITMistral Small 3
LMArena Expert10391202
GPQA Diamond—47.3%
Confabulations—25.2%

Multilingual Mistral Small 3 leads

Gemma 1.1 7b IT: 28.1 (#273), Mistral Small 3: 37.3 (#236)

Multilingual benchmarks
BenchmarkGemma 1.1 7b ITMistral Small 3
LMArena Non-English10521198
LMArena Chinese10611204
LMArena French10651203
LMArena German10541211
LMArena Japanese9711111
LMArena Korean9881188
LMArena Russian10461216
LMArena Spanish1049—

Instruction Following Mistral Small 3 leads

Gemma 1.1 7b IT: 54.0 (#283), Mistral Small 3: 63.7 (#229)

Instruction Following benchmarks
BenchmarkGemma 1.1 7b ITMistral Small 3
LMArena Instruction Following10571214

Long Context Mistral Small 3 leads

Gemma 1.1 7b IT: 32.1 (#272), Mistral Small 3: 37.8 (#211)

Long Context benchmarks
BenchmarkGemma 1.1 7b ITMistral Small 3
LMArena Longer Query10561246

Writing & Preference Mistral Small 3 leads

Gemma 1.1 7b IT: 30.4 (#288), Mistral Small 3: 32.2 (#280)

Writing & Preference benchmarks
BenchmarkGemma 1.1 7b ITMistral Small 3
LMArena Text10941234
LMArena Creative Writing10601195
LMArena Multi-Turn10401217
EQ-Bench Creative Writing—707

Frequently asked questions

Is Gemma 1.1 7b IT better than Mistral Small 3?

Gemma 1.1 7b IT and Mistral Small 3 score almost the same on the Noometry Index (31.3 vs 31.2), so choose on price, context window or the category you care about most.

Is Gemma 1.1 7b IT or Mistral Small 3 better for coding?

Mistral Small 3 scores higher on coding benchmarks: 36.5 versus 31.5 in the Noometry coding category.

How many benchmarks do Gemma 1.1 7b IT and Mistral Small 3 share?

16 benchmarks have published results for both models. Gemma 1.1 7b IT has 19 scored results on Noometry and Mistral Small 3 has 24.

Related comparisons

Go deeper