Model comparison

Mistral Small 3 vs Mistral Small 3.2

Mistral Small 3 and Mistral Small 3.2 score almost the same on the Noometry Index (31.2 vs 31.2), so choose on price, context window or the category you care about most.

Last verified . 5 shared benchmarks.

Mistral Small 3 Mistral AI

31.2

Rank #278 Confirmed

Mistral Small 3.2 Mistral AI

31.2

Rank #280 Confirmed

Summary

  • They share 5 benchmarks with published results for both. Mistral Small 3 scores higher in 1 category and Mistral Small 3.2 in 3 categories; 3 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Mistral Small 3.2 leads 45.0 to 32.2.
  • The biggest single-benchmark swing is OTIS Mock AIME 2024-2025: 6.7% for Mistral Small 3 and 30.3% for Mistral Small 3.2.
  • Mistral Small 3 is cheaper at $0.05 / $0.08 per million input/output tokens, against $0.0938 / $0.25 for Mistral Small 3.2.
  • Mistral Small 3.2 accepts more context: 256K tokens versus 33K.

Side by side

Mistral Small 3 and Mistral Small 3.2 specifications
Mistral Small 3Mistral Small 3.2
ProviderMistral AIMistral AI
Noometry Index31.231.2
Released2025-01-302025-06-20
WeightsOpenOpen
Context window33K256K
Max output16K16K
Input $ / M tokens$0.05$0.0938
Output $ / M tokens$0.08$0.25
Results tracked246

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Mistral Small 3: 36.5 (#207), Mistral Small 3.2: —

Coding benchmarks
BenchmarkMistral Small 3Mistral Small 3.2
BigCodeBench Instruct45.3%—
LMArena Coding1246—
BigCodeBench Complete50.4%—

Reasoning Too close to call

Mistral Small 3: 18.9 (#273), Mistral Small 3.2: 18.1 (#287)

Reasoning benchmarks
BenchmarkMistral Small 3Mistral Small 3.2
Chess Puzzles0%1%
Epoch Capabilities Index127.07131.74
Kagi LLM Benchmark—40.4%
LMArena Hard Prompts1233—

Math Mistral Small 3.2 leads

Mistral Small 3: 16.3 (#295), Mistral Small 3.2: 26.3 (#260)

Math benchmarks
BenchmarkMistral Small 3Mistral Small 3.2
OTIS Mock AIME 2024-20256.7%30.3%
LMArena Math1240—

Knowledge Mistral Small 3.2 leads

Mistral Small 3: 25.1 (#263), Mistral Small 3.2: 26.7 (#256)

Knowledge benchmarks
BenchmarkMistral Small 3Mistral Small 3.2
GPQA Diamond47.3%49.1%
Confabulations25.2%—
LMArena Expert1202—

Multilingual Not comparable

Mistral Small 3: 37.3 (#236), Mistral Small 3.2: —

Multilingual benchmarks
BenchmarkMistral Small 3Mistral Small 3.2
LMArena Non-English1198—
LMArena Chinese1204—
LMArena French1203—
LMArena German1211—
LMArena Japanese1111—
LMArena Korean1188—
LMArena Russian1216—

Instruction Following Not comparable

Mistral Small 3: 63.7 (#229), Mistral Small 3.2: —

Instruction Following benchmarks
BenchmarkMistral Small 3Mistral Small 3.2
LMArena Instruction Following1214—

Long Context Not comparable

Mistral Small 3: 37.8 (#211), Mistral Small 3.2: —

Long Context benchmarks
BenchmarkMistral Small 3Mistral Small 3.2
LMArena Longer Query1246—

Writing & Preference Mistral Small 3.2 leads

Mistral Small 3: 32.2 (#280), Mistral Small 3.2: 45.0 (#224)

Writing & Preference benchmarks
BenchmarkMistral Small 3Mistral Small 3.2
EQ-Bench Creative Writing7071255
LMArena Text1234—
LMArena Creative Writing1195—
LMArena Multi-Turn1217—

Frequently asked questions

Is Mistral Small 3 better than Mistral Small 3.2?

Mistral Small 3 and Mistral Small 3.2 score almost the same on the Noometry Index (31.2 vs 31.2), so choose on price, context window or the category you care about most.

Which is cheaper, Mistral Small 3 or Mistral Small 3.2?

Mistral Small 3 is cheaper. It lists at $0.05 per million input tokens and $0.08 per million output tokens; Mistral Small 3.2 lists at $0.0938 and $0.25.

Which has the bigger context window?

Mistral Small 3.2 does, with 256K tokens against 33K.

How many benchmarks do Mistral Small 3 and Mistral Small 3.2 share?

5 benchmarks have published results for both models. Mistral Small 3 has 24 scored results on Noometry and Mistral Small 3.2 has 6.

Related comparisons

Go deeper