Model comparison

Command R vs Pixtral Large

Command R and Pixtral Large score almost the same on the Noometry Index (31.4 vs 32.2), so choose on price, context window or the category you care about most.

Last verified . 0 shared benchmarks.

Command R Cohere

31.4

Rank #272 Confirmed

Pixtral Large Mistral AI

32.2

Rank #259 Reported

Summary

  • The widest gap is in reasoning, where Pixtral Large leads 21.7 to 13.8.
  • Command R is cheaper at $0.15 / $0.60 per million input/output tokens, against $2 / $6 for Pixtral Large.

Side by side

Command R and Pixtral Large specifications
Command RPixtral Large
ProviderCohereMistral AI
Noometry Index31.432.2
Released2024-08-302024-11-01
WeightsOpenOpen
Context window128K128K
Max output4K128K
Input $ / M tokens$0.15$2
Output $ / M tokens$0.60$6
Results tracked293

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Command R: 29.3 (#306), Pixtral Large: —

Coding benchmarks
BenchmarkCommand RPixtral Large
BigCodeBench Instruct37.1%—
LiveBench Coding17.9%—
LMArena Coding1169—
BigCodeBench Complete45.2%—

Reasoning Pixtral Large leads

Command R: 13.8 (#331), Pixtral Large: 21.7 (#218)

Reasoning benchmarks
BenchmarkCommand RPixtral Large
EnigmaEval—0.8%
LiveBench Reasoning21.9%—
LMArena Hard Prompts1164—
DTBench46.4%—
LiveBench Data Analysis33.3%—
LMCA9.2%—
LiveBench27.5%—

Math Not comparable

Command R: 28.0 (#246), Pixtral Large: —

Math benchmarks
BenchmarkCommand RPixtral Large
LiveBench Math19.4%—
LMArena Math1155—

Knowledge Not comparable

Command R: 31.0 (#221), Pixtral Large: —

Knowledge benchmarks
BenchmarkCommand RPixtral Large
LMArena Expert1138—
MMLU65.2%—

Multimodal Not comparable

Command R: —, Pixtral Large: 30.6 (#111)

Multimodal benchmarks
BenchmarkCommand RPixtral Large
LMArena Vision—1089

Multilingual Not comparable

Command R: 35.7 (#245), Pixtral Large: —

Multilingual benchmarks
BenchmarkCommand RPixtral Large
LMArena Non-English1174—
LMArena Chinese1182—
LMArena French1162—
LMArena German1176—
LMArena Japanese1143—
LMArena Korean1163—
LMArena Russian1174—
LMArena Spanish1151—

Instruction Following Not comparable

Command R: 58.1 (#261), Pixtral Large: —

Instruction Following benchmarks
BenchmarkCommand RPixtral Large
LiveBench Instruction Following55.6%—
LMArena Instruction Following1167—

Long Context Not comparable

Command R: 36.3 (#231), Pixtral Large: —

Long Context benchmarks
BenchmarkCommand RPixtral Large
LMArena Longer Query1198—

Writing & Preference Command R leads

Command R: 38.2 (#254), Pixtral Large: 32.9 (#278)

Writing & Preference benchmarks
BenchmarkCommand RPixtral Large
LMArena Text1187—
LMArena Creative Writing1170—
EQ-Bench Creative Writing—988
LMArena Multi-Turn1163—
LiveBench Language16.7%—

Frequently asked questions

Is Command R better than Pixtral Large?

Command R and Pixtral Large score almost the same on the Noometry Index (31.4 vs 32.2), so choose on price, context window or the category you care about most.

Which is cheaper, Command R or Pixtral Large?

Command R is cheaper. It lists at $0.15 per million input tokens and $0.60 per million output tokens; Pixtral Large lists at $2 and $6.

Which has the bigger context window?

Both accept 128K tokens.

How many benchmarks do Command R and Pixtral Large share?

0 benchmarks have published results for both models. Command R has 29 scored results on Noometry and Pixtral Large has 3.

Related comparisons

Go deeper