Model comparison

Command R vs Mistral Medium 3.1

Command R and Mistral Medium 3.1 score almost the same on the Noometry Index (31.4 vs 31.9), so choose on price, context window or the category you care about most.

Last verified . 0 shared benchmarks.

Command R Cohere

31.4

Rank #272 Confirmed

Mistral Medium 3.1 Mistral AI

31.9

Rank #266 Reported

Summary

  • The widest gap is in writing & preference, where Mistral Medium 3.1 leads 55.5 to 38.2.
  • Command R is cheaper at $0.15 / $0.60 per million input/output tokens, against $0.40 / $2 for Mistral Medium 3.1.
  • Mistral Medium 3.1 accepts more context: 131K tokens versus 128K.
  • Command R has downloadable open weights; the other is API-only.

Side by side

Command R and Mistral Medium 3.1 specifications
Command RMistral Medium 3.1
ProviderCohereMistral AI
Noometry Index31.431.9
Released2024-08-30—
WeightsOpenProprietary
Context window128K131K
Max output4K105K
Input $ / M tokens$0.15$0.40
Output $ / M tokens$0.60$2
Results tracked293

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Command R: 29.3 (#306), Mistral Medium 3.1: —

Coding benchmarks
BenchmarkCommand RMistral Medium 3.1
BigCodeBench Instruct37.1%—
LiveBench Coding17.9%—
LMArena Coding1169—
BigCodeBench Complete45.2%—

Reasoning Command R leads

Command R: 13.8 (#331), Mistral Medium 3.1: 10.6 (#341)

Reasoning benchmarks
BenchmarkCommand RMistral Medium 3.1
NYT Connections (extended)—6.5%
Thematic Generalization—20.3%
LiveBench Reasoning21.9%—
LMArena Hard Prompts1164—
DTBench46.4%—
LiveBench Data Analysis33.3%—
LMCA9.2%—
LiveBench27.5%—

Math Not comparable

Command R: 28.0 (#246), Mistral Medium 3.1: —

Math benchmarks
BenchmarkCommand RMistral Medium 3.1
LiveBench Math19.4%—
LMArena Math1155—

Knowledge Not comparable

Command R: 31.0 (#221), Mistral Medium 3.1: —

Knowledge benchmarks
BenchmarkCommand RMistral Medium 3.1
LMArena Expert1138—
MMLU65.2%—

Multilingual Not comparable

Command R: 35.7 (#245), Mistral Medium 3.1: —

Multilingual benchmarks
BenchmarkCommand RMistral Medium 3.1
LMArena Non-English1174—
LMArena Chinese1182—
LMArena French1162—
LMArena German1176—
LMArena Japanese1143—
LMArena Korean1163—
LMArena Russian1174—
LMArena Spanish1151—

Instruction Following Not comparable

Command R: 58.1 (#261), Mistral Medium 3.1: —

Instruction Following benchmarks
BenchmarkCommand RMistral Medium 3.1
LiveBench Instruction Following55.6%—
LMArena Instruction Following1167—

Long Context Not comparable

Command R: 36.3 (#231), Mistral Medium 3.1: —

Long Context benchmarks
BenchmarkCommand RMistral Medium 3.1
LMArena Longer Query1198—

Writing & Preference Mistral Medium 3.1 leads

Command R: 38.2 (#254), Mistral Medium 3.1: 55.5 (#145)

Writing & Preference benchmarks
BenchmarkCommand RMistral Medium 3.1
LMArena Text1187—
LMArena Creative Writing1170—
EQ-Bench Creative Writing—1476
LMArena Multi-Turn1163—
LiveBench Language16.7%—

Frequently asked questions

Is Command R better than Mistral Medium 3.1?

Command R and Mistral Medium 3.1 score almost the same on the Noometry Index (31.4 vs 31.9), so choose on price, context window or the category you care about most.

Which is cheaper, Command R or Mistral Medium 3.1?

Command R is cheaper. It lists at $0.15 per million input tokens and $0.60 per million output tokens; Mistral Medium 3.1 lists at $0.40 and $2.

Which has the bigger context window?

Mistral Medium 3.1 does, with 131K tokens against 128K.

How many benchmarks do Command R and Mistral Medium 3.1 share?

0 benchmarks have published results for both models. Command R has 29 scored results on Noometry and Mistral Medium 3.1 has 3.

Related comparisons

Go deeper