Google, open weights

Gemma 2 9B

Gemma 2 9B by Google ranks 341st of 354 ranked models on the Noometry Index as of October 2026, with a score of 25.9. Its strongest category is long context, where it ranks 233rd.

Last verified

Specifications

Noometry rank
#341 of 354
Index score
25.9
Evidence
Confirmed 35 results
Provider
Google
Released
June 24, 2024
Weights
Open weights
Reasoning
Unknown
Context window
—
Max output
—
Input price
Not listed
Output price
Not listed
Blended price
Not listed
Output speed
Not measured
Value
Not ranked
Knowledge cutoff
Unknown

Category scores

Each category score combines every public result we have in that category.

Gemma 2 9B category scores
  1. Coding 29.4
  2. Reasoning 15.9
  3. Math 9.9
  4. Knowledge 9.7
  5. Multilingual 36.6
  6. Instruction Following 57.6
  7. Long Context 36.3
  8. Writing & Preference 32.1
Gemma 2 9B category ranks
CategoryScoreRankResults
Coding29.4#3044
Reasoning15.9#3093
Math9.9#3184
Knowledge9.7#3052
Multilingual36.6#2381
Instruction Following57.6#2692
Long Context36.3#2331
Writing & Preference32.1#2815

Strengths and weaknesses

Categories where Gemma 2 9B places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

Gemma 2 9B: strongest categories
CategoryScorevs medianRank
Long Context36.3−4.7#233 of 296, top 79%
Multilingual36.6−10.8#238 of 297, top 81%
Instruction Following57.6−13.7#269 of 305, top 89%

Weakest categories

Gemma 2 9B: weakest categories
CategoryScorevs medianRank
Math9.9−26.6#318 of 327, top 98%
Knowledge9.7−27.6#305 of 314, top 98%
Writing & Preference32.1−21.6#281 of 312, top 91%

Closest competitors

The models ranked just above and below Gemma 2 9B. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Gemma 2 9B
ModelRankScoreBlended $/MSpeed
Mistral Nemo#33726.4$0.15—Compare
Ministral 3B#33826.2$0.10—Compare
DeepSeek-R1-Distill-Qwen-1.5B#33926.1——Compare
Claude 3 Haiku#34025.9—41Compare
Dolly 2.0-12b#34225.5——Compare
GPT-4o mini#34325.5$0.26120Compare
Llama 3-8B#34425.5——Compare
Claude 2.1#34525.2——Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

Gemma 2 9B Coding benchmark results
BenchmarkScorePositionSettingSourceDate
BigCodeBench Instruct34.7%#47 of 64, top 74%BigCodeBench2024-06-19
LiveBench Coding22.5%#33 of 39, top 85%Epoch AI
LMArena Coding1173#249 of 294, top 85%LMArena2026-10-08
BigCodeBench Complete40.6%#50 of 66, top 76%BigCodeBench2024-06-19

Reasoning

Gemma 2 9B Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
LiveBench Reasoning15.2%#39 of 39, top 100%Epoch AI
LMArena Hard Prompts1171#248 of 297, top 84%LMArena2026-10-08
LiveBench Data Analysis36.4%#34 of 39, top 88%Epoch AI
Epoch Capabilities Index119.83#169 of 213, top 80%Epoch AI2024-06-24
LiveBench28.7%#35 of 39, top 90%Epoch AI
PIQA83.7%#8 of 27, top 30%Epoch AI

Math

Gemma 2 9B Math benchmark results
BenchmarkScorePositionSettingSourceDate
OTIS Mock AIME 2024-20250.6%#170 of 173, top 99%Epoch AI2025-03-07
LiveBench Math19.8%#35 of 39, top 90%Epoch AI
LMArena Math1183#239 of 285, top 84%LMArena2026-10-08
MATH Level 521%#63 of 79, top 80%Epoch AI2025-01-27
GSM8K84.9%#9 of 38, top 24%Epoch AI

Knowledge

Gemma 2 9B Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
GPQA Diamond27.5%#175 of 186, top 95%Epoch AI2025-01-27
LMArena Expert1147#237 of 273, top 87%LMArena2026-10-08
BoolQ85.7%#8 of 23, top 35%Epoch AI
MMLU72.1%#40 of 81, top 50%Epoch AI

Multilingual

Gemma 2 9B Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1188#238 of 297, top 81%LMArena2026-10-08
LMArena Chinese1185#238 of 285, top 84%LMArena2026-10-08
LMArena French1190#195 of 223, top 88%LMArena2026-10-08
LMArena German1186#196 of 231, top 85%LMArena2026-10-08
LMArena Japanese1144#177 of 211, top 84%LMArena2026-10-08
LMArena Korean1137#185 of 213, top 87%LMArena2026-10-08
LMArena Russian1200#234 of 283, top 83%LMArena2026-10-08
LMArena Spanish1200#194 of 226, top 86%LMArena2026-10-08

Instruction Following

Gemma 2 9B Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
LiveBench Instruction Following52.6%#36 of 39, top 93%Epoch AI
LMArena Instruction Following1178#244 of 298, top 82%LMArena2026-10-08

Long Context

Gemma 2 9B Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Longer Query1197#242 of 291, top 84%LMArena2026-10-08

Writing & Preference

Gemma 2 9B Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1207#242 of 297, top 82%LMArena2026-10-08
LMArena Creative Writing1206#229 of 295, top 78%LMArena2026-10-08
EQ-Bench Creative Writing841#103 of 115, top 90%EQ-Bench
LMArena Multi-Turn1193#239 of 295, top 82%LMArena2026-10-08
LiveBench Language25.5%#32 of 39, top 83%Epoch AI

Compare Gemma 2 9B

Other Google models

Frequently asked questions

How good is Gemma 2 9B?

Gemma 2 9B by Google ranks 341st of 354 ranked models on the Noometry Index as of October 2026, with a score of 25.9. Its strongest category is long context, where it ranks 233rd.

Is Gemma 2 9B open source?

Yes. Gemma 2 9B's weights are downloadable; check the license for commercial terms.

What are Gemma 2 9B's strengths and weaknesses?

Relative to other ranked models, Gemma 2 9B places best in long context, multilingual, instruction following and lowest in math, knowledge, writing & preference.

What is Gemma 2 9B best at?

Its best category is long context, where it ranks 233rd on Noometry.