Model comparison

Command A vs Magistral Medium

Command A is the stronger model overall, scoring 36.5 to 35.2 on the Noometry Index. Magistral Medium costs 1.6× less per token, which makes it the better buy when Command A's lead doesn't matter for your workload.

Last verified . 18 shared benchmarks.

Command A Cohere

36.5

Rank #215 Confirmed

Magistral Medium Mistral AI

35.2

Rank #227 Confirmed

Summary

  • They share 18 benchmarks with published results for both. Command A scores higher in 7 categories and Magistral Medium in 1 category; 8 gaps are clear of the uncertainty.
  • The widest gap is in coding, where Magistral Medium leads 39.1 to 27.2.
  • The biggest single-benchmark swing is Kagi LLM Benchmark: 28.8% for Command A and 16.2% for Magistral Medium.
  • Magistral Medium is cheaper at $2 / $5 per million input/output tokens, against $2.50 / $10 for Command A.
  • Magistral Medium accepts more context: 262K tokens versus 256K.

Side by side

Command A and Magistral Medium specifications
Command AMagistral Medium
ProviderCohereMistral AI
Noometry Index36.535.2
Released2025-03-132025-03-17
WeightsOpenOpen
Context window256K262K
Max output8K16K
Input $ / M tokens$2.50$2
Output $ / M tokens$10$5
Results tracked2422

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Magistral Medium leads

Command A: 27.2 (#322), Magistral Medium: 39.1 (#161)

Coding benchmarks
BenchmarkCommand AMagistral Medium
LMArena Coding13301319
Aider Polyglot12%—
SciCode—39.2%

Agentic & Tool Use Not comparable

Command A: 35.9 (#40), Magistral Medium: —

Agentic & Tool Use benchmarks
BenchmarkCommand AMagistral Medium
Berkeley Function Calling Leaderboard57.1%—

Reasoning Command A leads

Command A: 18.3 (#283), Magistral Medium: 8.6 (#348)

Reasoning benchmarks
BenchmarkCommand AMagistral Medium
Kagi LLM Benchmark28.8%16.2%
LMArena Hard Prompts13261267
ARC-AGI-2—0%
ARC-AGI-1—6.1%
CritPt—0.3%
DTBench61.3%—
LMCA10.3%—

Math Command A leads

Command A: 36.2 (#171), Magistral Medium: 35.1 (#189)

Math benchmarks
BenchmarkCommand AMagistral Medium
LMArena Math13001250

Knowledge Command A leads

Command A: 37.1 (#159), Magistral Medium: 33.5 (#202)

Knowledge benchmarks
BenchmarkCommand AMagistral Medium
LMArena Expert12951223
Vectara Hallucination Rate9.3%—

Multilingual Command A leads

Command A: 45.3 (#170), Magistral Medium: 39.6 (#224)

Multilingual benchmarks
BenchmarkCommand AMagistral Medium
LMArena Non-English13131232
LMArena Chinese13271227
LMArena French13511267
LMArena German13411248
LMArena Japanese12851175
LMArena Korean12851125
LMArena Russian13141224
LMArena Spanish13471271

Instruction Following Command A leads

Command A: 69.1 (#177), Magistral Medium: 66.0 (#211)

Instruction Following benchmarks
BenchmarkCommand AMagistral Medium
LMArena Instruction Following13091254

Long Context Command A leads

Command A: 40.6 (#151), Magistral Medium: 39.3 (#183)

Long Context benchmarks
BenchmarkCommand AMagistral Medium
LMArena Longer Query13341295

Writing & Preference Command A leads

Command A: 47.6 (#208), Magistral Medium: 46.3 (#219)

Writing & Preference benchmarks
BenchmarkCommand AMagistral Medium
LMArena Text13311255
LMArena Creative Writing13191245
LMArena Multi-Turn13391275
EQ-Bench Creative Writing1145—

Frequently asked questions

Is Command A better than Magistral Medium?

Command A is the stronger model overall, scoring 36.5 to 35.2 on the Noometry Index. Magistral Medium costs 1.6× less per token, which makes it the better buy when Command A's lead doesn't matter for your workload.

Which is cheaper, Command A or Magistral Medium?

Magistral Medium is cheaper. It lists at $2 per million input tokens and $5 per million output tokens; Command A lists at $2.50 and $10.

Is Command A or Magistral Medium better for coding?

Magistral Medium scores higher on coding benchmarks: 39.1 versus 27.2 in the Noometry coding category.

Which has the bigger context window?

Magistral Medium does, with 262K tokens against 256K.

How many benchmarks do Command A and Magistral Medium share?

18 benchmarks have published results for both models. Command A has 24 scored results on Noometry and Magistral Medium has 22.

Related comparisons

Go deeper