Model comparison

Command R+ vs Gemini 1.5 Pro (May 2024)

Command R+ and Gemini 1.5 Pro (May 2024) score almost the same on the Noometry Index (32.4 vs 32.1), so choose on price, context window or the category you care about most.

Last verified . 25 shared benchmarks.

Command R+ Cohere

32.4

Rank #257 Confirmed

Gemini 1.5 Pro (May 2024) Google

32.1

Rank #261 Confirmed

Summary

  • They share 25 benchmarks with published results for both. Command R+ scores higher in 2 categories and Gemini 1.5 Pro (May 2024) in 6 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Gemini 1.5 Pro (May 2024) leads 52.4 to 43.5.
  • The biggest single-benchmark swing is BigCodeBench Complete: 41.9% for Command R+ and 57.5% for Gemini 1.5 Pro (May 2024).
  • Command R+ has downloadable open weights; the other is API-only.

Side by side

Command R+ and Gemini 1.5 Pro (May 2024) specifications
Command R+Gemini 1.5 Pro (May 2024)
ProviderCohereGoogle
Noometry Index32.432.1
Released2024-08-302024-02-15
WeightsOpenProprietary
Context window128K—
Max output4K—
Input $ / M tokens$2.50—
Output $ / M tokens$10—
Results tracked3445

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Gemini 1.5 Pro (May 2024) leads

Command R+: 29.1 (#309), Gemini 1.5 Pro (May 2024): 34.2 (#241)

Coding benchmarks
BenchmarkCommand R+Gemini 1.5 Pro (May 2024)
BigCodeBench Instruct33.8%43.8%
LMArena Coding11871294
BigCodeBench Complete41.9%57.5%
HumanEval+56.7%79.3%
MBPP+63.5%74.6%
WeirdML—22.2%
LiveBench Coding19.1%—
CadEval—34%

Agentic & Tool Use Not comparable

Command R+: —, Gemini 1.5 Pro (May 2024): 17.9 (#145)

Agentic & Tool Use benchmarks
BenchmarkCommand R+Gemini 1.5 Pro (May 2024)
TheAgentCompany—3.4%
Cybench—7.5%
BALROG—21%

Reasoning Gemini 1.5 Pro (May 2024) leads

Command R+: 9.2 (#344), Gemini 1.5 Pro (May 2024): 12.3 (#338)

Reasoning benchmarks
BenchmarkCommand R+Gemini 1.5 Pro (May 2024)
SimpleBench17.4%27.1%
LMArena Hard Prompts11861296
DTBench54.9%59%
Epoch Capabilities Index119.34131.73
ARC-AGI-2—0.8%
LiveBench Reasoning24.8%—
LiveBench Data Analysis38.1%—
LMCA5%—
BIG-Bench Hard—89.2%
ForecastBench—58.4
LiveBench31.8%—

Math Command R+ leads

Command R+: 28.9 (#242), Gemini 1.5 Pro (May 2024): 25.8 (#266)

Math benchmarks
BenchmarkCommand R+Gemini 1.5 Pro (May 2024)
LMArena Math11881315
OTIS Mock AIME 2024-2025—23.1%
Omni-MATH—36.4%
LiveBench Math21.3%—
MATH Level 5—70.4%

Knowledge Command R+ leads

Command R+: 36.4 (#169), Gemini 1.5 Pro (May 2024): 29.4 (#239)

Knowledge benchmarks
BenchmarkCommand R+Gemini 1.5 Pro (May 2024)
LMArena Expert11741279
MMLU69.4%86.9%
GPQA Diamond—57.2%
Humanity's Last Exam—4.6%
MMLU-Pro—73.7%
Confabulations—13.5%
Vectara Hallucination Rate6.9%—
GPQA (HELM)—53.4%

Multimodal Not comparable

Command R+: —, Gemini 1.5 Pro (May 2024): 36.8 (#77)

Multimodal benchmarks
BenchmarkCommand R+Gemini 1.5 Pro (May 2024)
LMArena Vision—1161
Video-MME—75%

Multilingual Gemini 1.5 Pro (May 2024) leads

Command R+: 38.6 (#227), Gemini 1.5 Pro (May 2024): 45.3 (#174)

Multilingual benchmarks
BenchmarkCommand R+Gemini 1.5 Pro (May 2024)
LMArena Non-English12161312
LMArena Chinese12261331
LMArena French12091302
LMArena German12161286
LMArena Japanese11661292
LMArena Korean11381298
LMArena Russian12271320
LMArena Spanish11891311

Instruction Following Gemini 1.5 Pro (May 2024) leads

Command R+: 60.0 (#254), Gemini 1.5 Pro (May 2024): 68.6 (#185)

Instruction Following benchmarks
BenchmarkCommand R+Gemini 1.5 Pro (May 2024)
LMArena Instruction Following11971297
LiveBench Instruction Following57.6%—
IFEval—83.7%

Long Context Gemini 1.5 Pro (May 2024) leads

Command R+: 37.3 (#219), Gemini 1.5 Pro (May 2024): 39.8 (#169)

Long Context benchmarks
BenchmarkCommand R+Gemini 1.5 Pro (May 2024)
LMArena Longer Query12301308

Writing & Preference Gemini 1.5 Pro (May 2024) leads

Command R+: 43.5 (#228), Gemini 1.5 Pro (May 2024): 52.4 (#172)

Writing & Preference benchmarks
BenchmarkCommand R+Gemini 1.5 Pro (May 2024)
LMArena Text12291319
LMArena Creative Writing12351333
LMArena Multi-Turn12131296
WildBench—81.3%
LiveBench Language29.7%—

Frequently asked questions

Is Command R+ better than Gemini 1.5 Pro (May 2024)?

Command R+ and Gemini 1.5 Pro (May 2024) score almost the same on the Noometry Index (32.4 vs 32.1), so choose on price, context window or the category you care about most.

Is Command R+ or Gemini 1.5 Pro (May 2024) better for coding?

Gemini 1.5 Pro (May 2024) scores higher on coding benchmarks: 34.2 versus 29.1 in the Noometry coding category.

How many benchmarks do Command R+ and Gemini 1.5 Pro (May 2024) share?

25 benchmarks have published results for both models. Command R+ has 34 scored results on Noometry and Gemini 1.5 Pro (May 2024) has 45.

Related comparisons

Go deeper