Model comparison

Command R vs Gemma 2 2b IT

Gemma 2 2b IT is the stronger model overall, scoring 33.1 to 31.4 on the Noometry Index.

Last verified . 17 shared benchmarks.

Command R Cohere

31.4

Rank #272 Confirmed

Gemma 2 2b IT Google

33.1

Rank #248 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Command R scores higher in 5 categories and Gemma 2 2b IT in 3 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Gemma 2 2b IT leads 21.4 to 13.8.

Side by side

Command R and Gemma 2 2b IT specifications
Command RGemma 2 2b IT
ProviderCohereGoogle
Noometry Index31.433.1
Released2024-08-30—
WeightsOpenOpen
Context window128K—
Max output4K—
Input $ / M tokens$0.15—
Output $ / M tokens$0.60—
Results tracked2917

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Gemma 2 2b IT leads

Command R: 29.3 (#306), Gemma 2 2b IT: 32.3 (#273)

Coding benchmarks
BenchmarkCommand RGemma 2 2b IT
LMArena Coding11691112
BigCodeBench Instruct37.1%—
LiveBench Coding17.9%—
BigCodeBench Complete45.2%—

Reasoning Gemma 2 2b IT leads

Command R: 13.8 (#331), Gemma 2 2b IT: 21.4 (#224)

Reasoning benchmarks
BenchmarkCommand RGemma 2 2b IT
LMArena Hard Prompts11641113
LiveBench Reasoning21.9%—
DTBench46.4%—
LiveBench Data Analysis33.3%—
LMCA9.2%—
LiveBench27.5%—

Math Gemma 2 2b IT leads

Command R: 28.0 (#246), Gemma 2 2b IT: 32.6 (#212)

Math benchmarks
BenchmarkCommand RGemma 2 2b IT
LMArena Math11551135
LiveBench Math19.4%—

Knowledge Command R leads

Command R: 31.0 (#221), Gemma 2 2b IT: 29.9 (#231)

Knowledge benchmarks
BenchmarkCommand RGemma 2 2b IT
LMArena Expert11381096
MMLU65.2%—

Multilingual Command R leads

Command R: 35.7 (#245), Gemma 2 2b IT: 32.3 (#257)

Multilingual benchmarks
BenchmarkCommand RGemma 2 2b IT
LMArena Non-English11741121
LMArena Chinese11821132
LMArena French11621157
LMArena German11761114
LMArena Japanese11431083
LMArena Korean11631055
LMArena Russian11741118
LMArena Spanish11511140

Instruction Following Too close to call

Command R: 58.1 (#261), Gemma 2 2b IT: 57.8 (#263)

Instruction Following benchmarks
BenchmarkCommand RGemma 2 2b IT
LMArena Instruction Following11671118
LiveBench Instruction Following55.6%—

Long Context Command R leads

Command R: 36.3 (#231), Gemma 2 2b IT: 34.3 (#250)

Long Context benchmarks
BenchmarkCommand RGemma 2 2b IT
LMArena Longer Query11981130

Writing & Preference Command R leads

Command R: 38.2 (#254), Gemma 2 2b IT: 36.5 (#263)

Writing & Preference benchmarks
BenchmarkCommand RGemma 2 2b IT
LMArena Text11871156
LMArena Creative Writing11701147
LMArena Multi-Turn11631118
LiveBench Language16.7%—

Frequently asked questions

Is Command R better than Gemma 2 2b IT?

Gemma 2 2b IT is the stronger model overall, scoring 33.1 to 31.4 on the Noometry Index.

Is Command R or Gemma 2 2b IT better for coding?

Gemma 2 2b IT scores higher on coding benchmarks: 32.3 versus 29.3 in the Noometry coding category.

How many benchmarks do Command R and Gemma 2 2b IT share?

17 benchmarks have published results for both models. Command R has 29 scored results on Noometry and Gemma 2 2b IT has 17.

Related comparisons

Go deeper