Model comparison

Command R+ vs Pixtral Large

Command R+ and Pixtral Large score almost the same on the Noometry Index (32.4 vs 32.2), so choose on price, context window or the category you care about most.

Last verified . 0 shared benchmarks.

Command R+ Cohere

32.4

Rank #257 Confirmed

Pixtral Large Mistral AI

32.2

Rank #259 Reported

Summary

  • The widest gap is in reasoning, where Pixtral Large leads 21.7 to 9.2.
  • Pixtral Large is cheaper at $2 / $6 per million input/output tokens, against $2.50 / $10 for Command R+.

Side by side

Command R+ and Pixtral Large specifications
Command R+Pixtral Large
ProviderCohereMistral AI
Noometry Index32.432.2
Released2024-08-302024-11-01
WeightsOpenOpen
Context window128K128K
Max output4K128K
Input $ / M tokens$2.50$2
Output $ / M tokens$10$6
Results tracked343

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Command R+: 29.1 (#309), Pixtral Large: —

Coding benchmarks
BenchmarkCommand R+Pixtral Large
BigCodeBench Instruct33.8%—
LiveBench Coding19.1%—
LMArena Coding1187—
BigCodeBench Complete41.9%—
HumanEval+56.7%—
MBPP+63.5%—

Reasoning Pixtral Large leads

Command R+: 9.2 (#344), Pixtral Large: 21.7 (#218)

Reasoning benchmarks
BenchmarkCommand R+Pixtral Large
SimpleBench17.4%—
EnigmaEval—0.8%
LiveBench Reasoning24.8%—
LMArena Hard Prompts1186—
DTBench54.9%—
LiveBench Data Analysis38.1%—
LMCA5%—
Epoch Capabilities Index119.34—
LiveBench31.8%—

Math Not comparable

Command R+: 28.9 (#242), Pixtral Large: —

Math benchmarks
BenchmarkCommand R+Pixtral Large
LiveBench Math21.3%—
LMArena Math1188—

Knowledge Not comparable

Command R+: 36.4 (#169), Pixtral Large: —

Knowledge benchmarks
BenchmarkCommand R+Pixtral Large
Vectara Hallucination Rate6.9%—
LMArena Expert1174—
MMLU69.4%—

Multimodal Not comparable

Command R+: —, Pixtral Large: 30.6 (#111)

Multimodal benchmarks
BenchmarkCommand R+Pixtral Large
LMArena Vision—1089

Multilingual Not comparable

Command R+: 38.6 (#227), Pixtral Large: —

Multilingual benchmarks
BenchmarkCommand R+Pixtral Large
LMArena Non-English1216—
LMArena Chinese1226—
LMArena French1209—
LMArena German1216—
LMArena Japanese1166—
LMArena Korean1138—
LMArena Russian1227—
LMArena Spanish1189—

Instruction Following Not comparable

Command R+: 60.0 (#254), Pixtral Large: —

Instruction Following benchmarks
BenchmarkCommand R+Pixtral Large
LiveBench Instruction Following57.6%—
LMArena Instruction Following1197—

Long Context Not comparable

Command R+: 37.3 (#219), Pixtral Large: —

Long Context benchmarks
BenchmarkCommand R+Pixtral Large
LMArena Longer Query1230—

Writing & Preference Command R+ leads

Command R+: 43.5 (#228), Pixtral Large: 32.9 (#278)

Writing & Preference benchmarks
BenchmarkCommand R+Pixtral Large
LMArena Text1229—
LMArena Creative Writing1235—
EQ-Bench Creative Writing—988
LMArena Multi-Turn1213—
LiveBench Language29.7%—

Frequently asked questions

Is Command R+ better than Pixtral Large?

Command R+ and Pixtral Large score almost the same on the Noometry Index (32.4 vs 32.2), so choose on price, context window or the category you care about most.

Which is cheaper, Command R+ or Pixtral Large?

Pixtral Large is cheaper. It lists at $2 per million input tokens and $6 per million output tokens; Command R+ lists at $2.50 and $10.

Which has the bigger context window?

Both accept 128K tokens.

How many benchmarks do Command R+ and Pixtral Large share?

0 benchmarks have published results for both models. Command R+ has 34 scored results on Noometry and Pixtral Large has 3.

Related comparisons

Go deeper