Model comparison

Command R vs Gemini 4 Argon

Gemini 4 Argon is the stronger model overall, scoring 56.5 to 31.4 on the Noometry Index.

Last verified . 13 shared benchmarks.

Command R Cohere

31.4

Rank #272 Confirmed

Gemini 4 Argon Google

56.5

Rank #24 Confirmed

Summary

  • They share 13 benchmarks with published results for both. Command R scores higher in 0 categories and Gemini 4 Argon in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in math, where Gemini 4 Argon leads 73.9 to 28.0.
  • Command R has downloadable open weights; the other is API-only.

Side by side

Command R and Gemini 4 Argon specifications
Command RGemini 4 Argon
ProviderCohereGoogle
Noometry Index31.456.5
Released2024-08-302026-09-30
WeightsOpenProprietary
Context window128K—
Max output4K—
Input $ / M tokens$0.15—
Output $ / M tokens$0.60—
Results tracked2921

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Gemini 4 Argon leads

Command R: 29.3 (#306), Gemini 4 Argon: 62.0 (#9)

Coding benchmarks
BenchmarkCommand RGemini 4 Argon
LMArena Coding11691550
LMArena WebDev—1678
FrontierSWE—55%
SciCode—61.8%
BigCodeBench Instruct37.1%—
LiveBench Coding17.9%—
BigCodeBench Complete45.2%—

Agentic & Tool Use Not comparable

Command R: —, Gemini 4 Argon: 49.2 (#8)

Agentic & Tool Use benchmarks
BenchmarkCommand RGemini 4 Argon
APEX-Agents—82.2%
Vending-Bench 2—13,718

Reasoning Gemini 4 Argon leads

Command R: 13.8 (#331), Gemini 4 Argon: 44.2 (#47)

Reasoning benchmarks
BenchmarkCommand RGemini 4 Argon
LMArena Hard Prompts11641548
CritPt—27.1%
LiveBench Reasoning21.9%—
DTBench46.4%—
LiveBench Data Analysis33.3%—
LMCA9.2%—
LiveBench27.5%—

Math Gemini 4 Argon leads

Command R: 28.0 (#246), Gemini 4 Argon: 73.9 (#17)

Math benchmarks
BenchmarkCommand RGemini 4 Argon
LMArena Math11551536
ProofBench—99%
LiveBench Math19.4%—

Knowledge Gemini 4 Argon leads

Command R: 31.0 (#221), Gemini 4 Argon: 43.9 (#89)

Knowledge benchmarks
BenchmarkCommand RGemini 4 Argon
LMArena Expert11381553
MMLU65.2%—

Multimodal Not comparable

Command R: —, Gemini 4 Argon: 46.5 (#14)

Multimodal benchmarks
BenchmarkCommand RGemini 4 Argon
Blueprint-Bench 2—54.4%

Multilingual Gemini 4 Argon leads

Command R: 35.7 (#245), Gemini 4 Argon: 60.3 (#1)

Multilingual benchmarks
BenchmarkCommand RGemini 4 Argon
LMArena Non-English11741523
LMArena Chinese11821610
LMArena Russian11741537
LMArena Spanish11511497
LMArena French1162—
LMArena German1176—
LMArena Japanese1143—
LMArena Korean1163—

Instruction Following Gemini 4 Argon leads

Command R: 58.1 (#261), Gemini 4 Argon: 80.1 (#2)

Instruction Following benchmarks
BenchmarkCommand RGemini 4 Argon
LMArena Instruction Following11671538
LiveBench Instruction Following55.6%—

Long Context Gemini 4 Argon leads

Command R: 36.3 (#231), Gemini 4 Argon: 47.6 (#15)

Long Context benchmarks
BenchmarkCommand RGemini 4 Argon
LMArena Longer Query11981549

Writing & Preference Gemini 4 Argon leads

Command R: 38.2 (#254), Gemini 4 Argon: 71.4 (#19)

Writing & Preference benchmarks
BenchmarkCommand RGemini 4 Argon
LMArena Text11871534
LMArena Creative Writing11701531
LMArena Multi-Turn11631557
LiveBench Language16.7%—

Frequently asked questions

Is Command R better than Gemini 4 Argon?

Gemini 4 Argon is the stronger model overall, scoring 56.5 to 31.4 on the Noometry Index.

Is Command R or Gemini 4 Argon better for coding?

Gemini 4 Argon scores higher on coding benchmarks: 62.0 versus 29.3 in the Noometry coding category.

How many benchmarks do Command R and Gemini 4 Argon share?

13 benchmarks have published results for both models. Command R has 29 scored results on Noometry and Gemini 4 Argon has 21.

Related comparisons

Go deeper