Model comparison

Gemma 1.1 7b IT vs Phi-4 Mini

Gemma 1.1 7b IT and Phi-4 Mini score almost the same on the Noometry Index (31.3 vs 30.9), so choose on price, context window or the category you care about most.

Last verified . 0 shared benchmarks.

Gemma 1.1 7b IT Google

31.3

Rank #277 Confirmed

Phi-4 Mini Microsoft

30.9

Rank #283 Reported

Summary

  • The widest gap is in coding, where Gemma 1.1 7b IT leads 31.5 to 28.1.

Side by side

Gemma 1.1 7b IT and Phi-4 Mini specifications
Gemma 1.1 7b ITPhi-4 Mini
ProviderGoogleMicrosoft
Noometry Index31.330.9
Released—2024-12-11
WeightsOpenOpen
Context window—128K
Max output—4K
Input $ / M tokens—$0.075
Output $ / M tokens—$0.30
Results tracked193

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Gemma 1.1 7b IT leads

Gemma 1.1 7b IT: 31.5 (#284), Phi-4 Mini: 28.1 (#317)

Coding benchmarks
BenchmarkGemma 1.1 7b ITPhi-4 Mini
SciCode—10.8%
LMArena Coding1084—
HumanEval+35.4%—
MBPP+45%—

Reasoning Phi-4 Mini leads

Gemma 1.1 7b IT: 20.5 (#238), Phi-4 Mini: 22.4 (#195)

Reasoning benchmarks
BenchmarkGemma 1.1 7b ITPhi-4 Mini
CritPt—0%
LMArena Hard Prompts1071—

Math Not comparable

Gemma 1.1 7b IT: 32.0 (#220), Phi-4 Mini: —

Math benchmarks
BenchmarkGemma 1.1 7b ITPhi-4 Mini
LMArena Math1107—

Knowledge Gemma 1.1 7b IT leads

Gemma 1.1 7b IT: 28.3 (#247), Phi-4 Mini: 25.3 (#262)

Knowledge benchmarks
BenchmarkGemma 1.1 7b ITPhi-4 Mini
Vectara Hallucination Rate—23.5%
LMArena Expert1039—

Multilingual Not comparable

Gemma 1.1 7b IT: 28.1 (#273), Phi-4 Mini: —

Multilingual benchmarks
BenchmarkGemma 1.1 7b ITPhi-4 Mini
LMArena Non-English1052—
LMArena Chinese1061—
LMArena French1065—
LMArena German1054—
LMArena Japanese971—
LMArena Korean988—
LMArena Russian1046—
LMArena Spanish1049—

Instruction Following Not comparable

Gemma 1.1 7b IT: 54.0 (#283), Phi-4 Mini: —

Instruction Following benchmarks
BenchmarkGemma 1.1 7b ITPhi-4 Mini
LMArena Instruction Following1057—

Long Context Not comparable

Gemma 1.1 7b IT: 32.1 (#272), Phi-4 Mini: —

Long Context benchmarks
BenchmarkGemma 1.1 7b ITPhi-4 Mini
LMArena Longer Query1056—

Writing & Preference Not comparable

Gemma 1.1 7b IT: 30.4 (#288), Phi-4 Mini: —

Writing & Preference benchmarks
BenchmarkGemma 1.1 7b ITPhi-4 Mini
LMArena Text1094—
LMArena Creative Writing1060—
LMArena Multi-Turn1040—

Frequently asked questions

Is Gemma 1.1 7b IT better than Phi-4 Mini?

Gemma 1.1 7b IT and Phi-4 Mini score almost the same on the Noometry Index (31.3 vs 30.9), so choose on price, context window or the category you care about most.

Is Gemma 1.1 7b IT or Phi-4 Mini better for coding?

Gemma 1.1 7b IT scores higher on coding benchmarks: 31.5 versus 28.1 in the Noometry coding category.

How many benchmarks do Gemma 1.1 7b IT and Phi-4 Mini share?

0 benchmarks have published results for both models. Gemma 1.1 7b IT has 19 scored results on Noometry and Phi-4 Mini has 3.

Related comparisons

Go deeper