Model comparison

C4ai Aya Expanse 32b vs Mistral Medium

C4ai Aya Expanse 32b and Mistral Medium score almost the same on the Noometry Index (35.9 vs 36.3), so choose on price, context window or the category you care about most.

Last verified . 18 shared benchmarks.

C4ai Aya Expanse 32b Cohere

35.9

Rank #221 Confirmed

Mistral Medium Mistral AI

36.3

Rank #218 Confirmed

Summary

  • They share 18 benchmarks with published results for both. C4ai Aya Expanse 32b scores higher in 3 categories and Mistral Medium in 5 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Mistral Medium leads 60.0 to 42.2.
  • The biggest single-benchmark swing is Vectara Hallucination Rate: 10.9% for C4ai Aya Expanse 32b and 22.7% for Mistral Medium.
  • Mistral Medium accepts more context: 262K tokens versus 128K.

Side by side

C4ai Aya Expanse 32b and Mistral Medium specifications
C4ai Aya Expanse 32bMistral Medium
ProviderCohereMistral AI
Noometry Index35.936.3
Released2024-10-242023-12-11
WeightsOpenOpen
Context window128K262K
Max output4K262K
Input $ / M tokens—$1.50
Output $ / M tokens—$7.50
Results tracked1836

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

C4ai Aya Expanse 32b: 34.8 (#231), Mistral Medium: 34.2 (#243)

Coding benchmarks
BenchmarkC4ai Aya Expanse 32bMistral Medium
LMArena Coding11971434
FrontierCode—8%
SciCode—40.2%
WeirdML—43.7%
ALE-Bench—763.98

Agentic & Tool Use Not comparable

C4ai Aya Expanse 32b: —, Mistral Medium: 28.3 (#90)

Agentic & Tool Use benchmarks
BenchmarkC4ai Aya Expanse 32bMistral Medium
Berkeley Function Calling Leaderboard—37.7%

Reasoning Too close to call

C4ai Aya Expanse 32b: 23.3 (#180), Mistral Medium: 24.0 (#167)

Reasoning benchmarks
BenchmarkC4ai Aya Expanse 32bMistral Medium
LMArena Hard Prompts11931426
Kagi LLM Benchmark—50%
CritPt—0%
DTBench—75.5%
LMCA—26.1%
Surface Evolver Bench—26.9%

Math C4ai Aya Expanse 32b leads

C4ai Aya Expanse 32b: 34.0 (#197), Mistral Medium: 28.1 (#245)

Math benchmarks
BenchmarkC4ai Aya Expanse 32bMistral Medium
LMArena Math12001408
OTIS Mock AIME 2024-2025—32.2%
ProofBench—9%
MATH Level 5—81.6%
FrontierMath (Feb 2025 set)—0.3%

Knowledge C4ai Aya Expanse 32b leads

C4ai Aya Expanse 32b: 33.2 (#206), Mistral Medium: 25.0 (#265)

Knowledge benchmarks
BenchmarkC4ai Aya Expanse 32bMistral Medium
Vectara Hallucination Rate10.9%22.7%
LMArena Expert11821408
GPQA Diamond—59.5%
Humanity's Last Exam—4.5%

Multimodal Not comparable

C4ai Aya Expanse 32b: —, Mistral Medium: 35.3 (#88)

Multimodal benchmarks
BenchmarkC4ai Aya Expanse 32bMistral Medium
LMArena Vision—1172

Multilingual Mistral Medium leads

C4ai Aya Expanse 32b: 38.4 (#230), Mistral Medium: 52.1 (#91)

Multilingual benchmarks
BenchmarkC4ai Aya Expanse 32bMistral Medium
LMArena Non-English12131408
LMArena Chinese12111447
LMArena French12481459
LMArena German11991432
LMArena Japanese11631378
LMArena Korean11581380
LMArena Russian12271411
LMArena Spanish11931433

Instruction Following Mistral Medium leads

C4ai Aya Expanse 32b: 62.6 (#237), Mistral Medium: 73.7 (#116)

Instruction Following benchmarks
BenchmarkC4ai Aya Expanse 32bMistral Medium
LMArena Instruction Following11961398

Long Context Mistral Medium leads

C4ai Aya Expanse 32b: 37.2 (#220), Mistral Medium: 42.9 (#114)

Long Context benchmarks
BenchmarkC4ai Aya Expanse 32bMistral Medium
LMArena Longer Query12281406

Writing & Preference Mistral Medium leads

C4ai Aya Expanse 32b: 42.2 (#235), Mistral Medium: 60.0 (#103)

Writing & Preference benchmarks
BenchmarkC4ai Aya Expanse 32bMistral Medium
LMArena Text12241424
LMArena Creative Writing12001391
LMArena Multi-Turn11901418
Short-Story Creative Writing—77.3%

Frequently asked questions

Is C4ai Aya Expanse 32b better than Mistral Medium?

C4ai Aya Expanse 32b and Mistral Medium score almost the same on the Noometry Index (35.9 vs 36.3), so choose on price, context window or the category you care about most.

Is C4ai Aya Expanse 32b or Mistral Medium better for coding?

They score almost the same on coding (34.8 vs 34.2); test both on your own repository before choosing.

Which has the bigger context window?

Mistral Medium does, with 262K tokens against 128K.

How many benchmarks do C4ai Aya Expanse 32b and Mistral Medium share?

18 benchmarks have published results for both models. C4ai Aya Expanse 32b has 18 scored results on Noometry and Mistral Medium has 36.

Related comparisons

Go deeper