Model comparison

Magistral Medium vs Pixtral Large

Magistral Medium is the stronger model overall, scoring 35.2 to 32.2 on the Noometry Index.

Last verified . 0 shared benchmarks.

Magistral Medium Mistral AI

35.2

Rank #227 Confirmed

Pixtral Large Mistral AI

32.2

Rank #259 Reported

Summary

  • The widest gap is in writing & preference, where Magistral Medium leads 46.3 to 32.9.
  • Magistral Medium is cheaper at $2 / $5 per million input/output tokens, against $2 / $6 for Pixtral Large.
  • Magistral Medium accepts more context: 262K tokens versus 128K.

Side by side

Magistral Medium and Pixtral Large specifications
Magistral MediumPixtral Large
ProviderMistral AIMistral AI
Noometry Index35.232.2
Released2025-03-172024-11-01
WeightsOpenOpen
Context window262K128K
Max output16K128K
Input $ / M tokens$2$2
Output $ / M tokens$5$6
Results tracked223

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Magistral Medium: 39.1 (#161), Pixtral Large: —

Coding benchmarks
BenchmarkMagistral MediumPixtral Large
SciCode39.2%—
LMArena Coding1319—

Reasoning Pixtral Large leads

Magistral Medium: 8.6 (#348), Pixtral Large: 21.7 (#218)

Reasoning benchmarks
BenchmarkMagistral MediumPixtral Large
ARC-AGI-20%—
Kagi LLM Benchmark16.2%—
ARC-AGI-16.1%—
CritPt0.3%—
EnigmaEval—0.8%
LMArena Hard Prompts1267—

Math Not comparable

Magistral Medium: 35.1 (#189), Pixtral Large: —

Math benchmarks
BenchmarkMagistral MediumPixtral Large
LMArena Math1250—

Knowledge Not comparable

Magistral Medium: 33.5 (#202), Pixtral Large: —

Knowledge benchmarks
BenchmarkMagistral MediumPixtral Large
LMArena Expert1223—

Multimodal Not comparable

Magistral Medium: —, Pixtral Large: 30.6 (#111)

Multimodal benchmarks
BenchmarkMagistral MediumPixtral Large
LMArena Vision—1089

Multilingual Not comparable

Magistral Medium: 39.6 (#224), Pixtral Large: —

Multilingual benchmarks
BenchmarkMagistral MediumPixtral Large
LMArena Non-English1232—
LMArena Chinese1227—
LMArena French1267—
LMArena German1248—
LMArena Japanese1175—
LMArena Korean1125—
LMArena Russian1224—
LMArena Spanish1271—

Instruction Following Not comparable

Magistral Medium: 66.0 (#211), Pixtral Large: —

Instruction Following benchmarks
BenchmarkMagistral MediumPixtral Large
LMArena Instruction Following1254—

Long Context Not comparable

Magistral Medium: 39.3 (#183), Pixtral Large: —

Long Context benchmarks
BenchmarkMagistral MediumPixtral Large
LMArena Longer Query1295—

Writing & Preference Magistral Medium leads

Magistral Medium: 46.3 (#219), Pixtral Large: 32.9 (#278)

Writing & Preference benchmarks
BenchmarkMagistral MediumPixtral Large
LMArena Text1255—
LMArena Creative Writing1245—
EQ-Bench Creative Writing—988
LMArena Multi-Turn1275—

Frequently asked questions

Is Magistral Medium better than Pixtral Large?

Magistral Medium is the stronger model overall, scoring 35.2 to 32.2 on the Noometry Index.

Which is cheaper, Magistral Medium or Pixtral Large?

Magistral Medium is cheaper. It lists at $2 per million input tokens and $5 per million output tokens; Pixtral Large lists at $2 and $6.

Which has the bigger context window?

Magistral Medium does, with 262K tokens against 128K.

How many benchmarks do Magistral Medium and Pixtral Large share?

0 benchmarks have published results for both models. Magistral Medium has 22 scored results on Noometry and Pixtral Large has 3.

Related comparisons

Go deeper