Model comparison

Command R vs Granite 3.0 8b Instruct

Command R and Granite 3.0 8b Instruct score almost the same on the Noometry Index (31.4 vs 31.6), so choose on price, context window or the category you care about most.

Last verified . 14 shared benchmarks.

Command R Cohere

31.4

Rank #272 Confirmed

Granite 3.0 8b Instruct IBM

31.6

Rank #270 Confirmed

Summary

  • They share 14 benchmarks with published results for both. Command R scores higher in 5 categories and Granite 3.0 8b Instruct in 3 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in multilingual, where Command R leads 35.7 to 27.2.
  • The biggest single-benchmark swing is BigCodeBench Complete: 45.2% for Command R and 35.4% for Granite 3.0 8b Instruct.

Side by side

Command R and Granite 3.0 8b Instruct specifications
Command RGranite 3.0 8b Instruct
ProviderCohereIBM
Noometry Index31.431.6
Released2024-08-30—
WeightsOpenOpen
Context window128K—
Max output4K—
Input $ / M tokens$0.15—
Output $ / M tokens$0.60—
Results tracked2914

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Command R: 29.3 (#306), Granite 3.0 8b Instruct: 29.7 (#301)

Coding benchmarks
BenchmarkCommand RGranite 3.0 8b Instruct
BigCodeBench Instruct37.1%29.3%
LMArena Coding11691112
BigCodeBench Complete45.2%35.4%
LiveBench Coding17.9%—

Reasoning Granite 3.0 8b Instruct leads

Command R: 13.8 (#331), Granite 3.0 8b Instruct: 20.9 (#230)

Reasoning benchmarks
BenchmarkCommand RGranite 3.0 8b Instruct
LMArena Hard Prompts11641092
LiveBench Reasoning21.9%—
DTBench46.4%—
LiveBench Data Analysis33.3%—
LMCA9.2%—
LiveBench27.5%—

Math Granite 3.0 8b Instruct leads

Command R: 28.0 (#246), Granite 3.0 8b Instruct: 32.8 (#210)

Math benchmarks
BenchmarkCommand RGranite 3.0 8b Instruct
LMArena Math11551143
LiveBench Math19.4%—

Knowledge Command R leads

Command R: 31.0 (#221), Granite 3.0 8b Instruct: 29.6 (#236)

Knowledge benchmarks
BenchmarkCommand RGranite 3.0 8b Instruct
LMArena Expert11381087
MMLU65.2%—

Multilingual Command R leads

Command R: 35.7 (#245), Granite 3.0 8b Instruct: 27.2 (#276)

Multilingual benchmarks
BenchmarkCommand RGranite 3.0 8b Instruct
LMArena Non-English11741037
LMArena Chinese11821063
LMArena Russian11741060
LMArena French1162—
LMArena German1176—
LMArena Japanese1143—
LMArena Korean1163—
LMArena Spanish1151—

Instruction Following Command R leads

Command R: 58.1 (#261), Granite 3.0 8b Instruct: 56.0 (#276)

Instruction Following benchmarks
BenchmarkCommand RGranite 3.0 8b Instruct
LMArena Instruction Following11671088
LiveBench Instruction Following55.6%—

Long Context Command R leads

Command R: 36.3 (#231), Granite 3.0 8b Instruct: 34.0 (#252)

Long Context benchmarks
BenchmarkCommand RGranite 3.0 8b Instruct
LMArena Longer Query11981122

Writing & Preference Command R leads

Command R: 38.2 (#254), Granite 3.0 8b Instruct: 31.1 (#285)

Writing & Preference benchmarks
BenchmarkCommand RGranite 3.0 8b Instruct
LMArena Text11871096
LMArena Creative Writing11701071
LMArena Multi-Turn11631063
LiveBench Language16.7%—

Frequently asked questions

Is Command R better than Granite 3.0 8b Instruct?

Command R and Granite 3.0 8b Instruct score almost the same on the Noometry Index (31.4 vs 31.6), so choose on price, context window or the category you care about most.

Is Command R or Granite 3.0 8b Instruct better for coding?

They score almost the same on coding (29.3 vs 29.7); test both on your own repository before choosing.

How many benchmarks do Command R and Granite 3.0 8b Instruct share?

14 benchmarks have published results for both models. Command R has 29 scored results on Noometry and Granite 3.0 8b Instruct has 14.

Related comparisons

Go deeper