Model comparison

Magistral Small vs Mistral

Magistral Small and Mistral score almost the same on the Noometry Index (30.2 vs 29.9), so choose on price, context window or the category you care about most.

Last verified . 0 shared benchmarks.

Magistral Small Mistral AI

30.2

Rank #296 Confirmed

Mistral Mistral AI

29.9

Rank #303 Confirmed

Summary

  • The widest gap is in reasoning, where Mistral leads 22.2 to 6.8.
  • Magistral Small has downloadable open weights; the other is API-only.

Side by side

Magistral Small and Mistral specifications
Magistral SmallMistral
ProviderMistral AIMistral AI
Noometry Index30.229.9
Released2025-06-10—
WeightsOpenProprietary
Context window128K—
Max output40K—
Input $ / M tokens$0.50—
Output $ / M tokens$1.50—
Results tracked1022

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Magistral Small leads

Magistral Small: 38.4 (#176), Mistral: 33.8 (#250)

Coding benchmarks
BenchmarkMagistral SmallMistral
SciCode35.2%—
LMArena Coding—1162

Reasoning Mistral leads

Magistral Small: 6.8 (#350), Mistral: 22.2 (#200)

Reasoning benchmarks
BenchmarkMagistral SmallMistral
ARC-AGI-20%—
Kagi LLM Benchmark6.3%—
ARC-AGI-15%—
CritPt0.3%—
Chess Puzzles3%—
LMArena Hard Prompts—1149
DTBench61.3%—
Epoch Capabilities Index133.19—

Math Magistral Small leads

Magistral Small: 26.2 (#261), Mistral: 22.3 (#278)

Math benchmarks
BenchmarkMagistral SmallMistral
OTIS Mock AIME 2024-202530%—
Omni-MATH—7.2%
LMArena Math—1180

Knowledge Magistral Small leads

Magistral Small: 30.9 (#223), Mistral: 16.6 (#288)

Knowledge benchmarks
BenchmarkMagistral SmallMistral
GPQA Diamond56.1%—
MMLU-Pro—27.7%
GPQA (HELM)—30.3%
LMArena Expert—1125

Multilingual Not comparable

Magistral Small: —, Mistral: 32.8 (#254)

Multilingual benchmarks
BenchmarkMagistral SmallMistral
LMArena Non-English—1129
LMArena Chinese—1109
LMArena French—1180
LMArena German—1155
LMArena Japanese—1013
LMArena Korean—1032
LMArena Russian—1168
LMArena Spanish—1143

Instruction Following Not comparable

Magistral Small: —, Mistral: 52.6 (#288)

Instruction Following benchmarks
BenchmarkMagistral SmallMistral
IFEval—56.8%
LMArena Instruction Following—1152

Long Context Not comparable

Magistral Small: —, Mistral: 35.0 (#245)

Long Context benchmarks
BenchmarkMagistral SmallMistral
LMArena Longer Query—1153

Writing & Preference Not comparable

Magistral Small: —, Mistral: 37.0 (#260)

Writing & Preference benchmarks
BenchmarkMagistral SmallMistral
LMArena Text—1165
LMArena Creative Writing—1158
WildBench—66%
LMArena Multi-Turn—1147

Frequently asked questions

Is Magistral Small better than Mistral?

Magistral Small and Mistral score almost the same on the Noometry Index (30.2 vs 29.9), so choose on price, context window or the category you care about most.

Is Magistral Small or Mistral better for coding?

Magistral Small scores higher on coding benchmarks: 38.4 versus 33.8 in the Noometry coding category.

How many benchmarks do Magistral Small and Mistral share?

0 benchmarks have published results for both models. Magistral Small has 10 scored results on Noometry and Mistral has 22.

Related comparisons

Go deeper