Model comparison

Gemini 1.5 Flash (May 2024) vs Mistral Small

Gemini 1.5 Flash (May 2024) and Mistral Small score almost the same on the Noometry Index (33.2 vs 33.4), so choose on price, context window or the category you care about most.

Last verified . 25 shared benchmarks.

Gemini 1.5 Flash (May 2024) Google

33.2

Rank #246 Confirmed

Mistral Small Mistral AI

33.4

Rank #243 Confirmed

Summary

  • They share 25 benchmarks with published results for both. Gemini 1.5 Flash (May 2024) scores higher in 5 categories and Mistral Small in 5 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in math, where Gemini 1.5 Flash (May 2024) leads 22.1 to 16.4.
  • The biggest single-benchmark swing is DTBench: 53.8% for Gemini 1.5 Flash (May 2024) and 70.9% for Mistral Small.
  • Mistral Small has downloadable open weights; the other is API-only.

Side by side

Gemini 1.5 Flash (May 2024) and Mistral Small specifications
Gemini 1.5 Flash (May 2024)Mistral Small
ProviderGoogleMistral AI
Noometry Index33.233.4
Released2024-05-142024-02-26
WeightsProprietaryOpen
Context window—262K
Max output—256K
Input $ / M tokens—$0.15
Output $ / M tokens—$0.60
Results tracked4239

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Gemini 1.5 Flash (May 2024): 34.4 (#236), Mistral Small: 34.0 (#247)

Coding benchmarks
BenchmarkGemini 1.5 Flash (May 2024)Mistral Small
BigCodeBench Instruct43.5%36.1%
LMArena Coding12611362
BigCodeBench Complete55.1%46.6%
SciCode—26.5%
WeirdML24.9%—
LiveBench Coding—36.2%
ALE-Bench—497.62
HumanEval+75.6%—
MBPP+67.5%—

Agentic & Tool Use Mistral Small leads

Gemini 1.5 Flash (May 2024): 26.6 (#102), Mistral Small: 28.1 (#93)

Agentic & Tool Use benchmarks
BenchmarkGemini 1.5 Flash (May 2024)Mistral Small
Berkeley Function Calling Leaderboard—37.1%
BALROG14.6%—

Reasoning Gemini 1.5 Flash (May 2024) leads

Gemini 1.5 Flash (May 2024): 21.7 (#215), Mistral Small: 19.8 (#250)

Reasoning benchmarks
BenchmarkGemini 1.5 Flash (May 2024)Mistral Small
LMArena Hard Prompts12571335
DTBench53.8%70.9%
Kagi LLM Benchmark—37.8%
CritPt—0%
LiveBench Reasoning—44.8%
LiveBench Data Analysis—53.7%
LMCA—20.6%
Epoch Capabilities Index129.36—
ForecastBench53.9—
LiveBench—44%
PIQA87.5%—

Math Gemini 1.5 Flash (May 2024) leads

Gemini 1.5 Flash (May 2024): 22.1 (#281), Mistral Small: 16.4 (#293)

Math benchmarks
BenchmarkGemini 1.5 Flash (May 2024)Mistral Small
OTIS Mock AIME 2024-202516.3%5.8%
LMArena Math12691341
MATH Level 561.9%46.8%
Omni-MATH30.4%—
LiveBench Math—39.9%
FrontierMath (Feb 2025 set)0%—
GSM8K82.4%—

Knowledge Mistral Small leads

Gemini 1.5 Flash (May 2024): 26.2 (#260), Mistral Small: 31.0 (#222)

Knowledge benchmarks
BenchmarkGemini 1.5 Flash (May 2024)Mistral Small
GPQA Diamond47.3%47.5%
LMArena Expert12331291
MMLU77.9%68.7%
MMLU-Pro67.8%—
Vectara Hallucination Rate—5.1%
GPQA (HELM)43.7%—
BoolQ85.8%—

Multimodal Gemini 1.5 Flash (May 2024) leads

Gemini 1.5 Flash (May 2024): 36.0 (#81), Mistral Small: 33.5 (#96)

Multimodal benchmarks
BenchmarkGemini 1.5 Flash (May 2024)Mistral Small
LMArena Vision11411142
Video-MME70.3%—
GeoBench76%—

Multilingual Mistral Small leads

Gemini 1.5 Flash (May 2024): 42.9 (#189), Mistral Small: 45.5 (#169)

Multilingual benchmarks
BenchmarkGemini 1.5 Flash (May 2024)Mistral Small
LMArena Non-English12781315
LMArena Chinese12951340
LMArena French12581337
LMArena German12621340
LMArena Japanese12521275
LMArena Korean12211259
LMArena Russian12881324
LMArena Spanish12431346

Instruction Following Too close to call

Gemini 1.5 Flash (May 2024): 66.8 (#205), Mistral Small: 66.4 (#209)

Instruction Following benchmarks
BenchmarkGemini 1.5 Flash (May 2024)Mistral Small
LMArena Instruction Following12581310
LiveBench Instruction Following—63.7%
IFEval83.1%—

Long Context Mistral Small leads

Gemini 1.5 Flash (May 2024): 39.0 (#187), Mistral Small: 40.4 (#156)

Long Context benchmarks
BenchmarkGemini 1.5 Flash (May 2024)Mistral Small
LMArena Longer Query12841327

Writing & Preference Mistral Small leads

Gemini 1.5 Flash (May 2024): 48.7 (#196), Mistral Small: 52.5 (#171)

Writing & Preference benchmarks
BenchmarkGemini 1.5 Flash (May 2024)Mistral Small
LMArena Text12871338
LMArena Creative Writing12851305
LMArena Multi-Turn12531344
WildBench79.2%—
LiveBench Language—30.5%

Frequently asked questions

Is Gemini 1.5 Flash (May 2024) better than Mistral Small?

Gemini 1.5 Flash (May 2024) and Mistral Small score almost the same on the Noometry Index (33.2 vs 33.4), so choose on price, context window or the category you care about most.

Is Gemini 1.5 Flash (May 2024) or Mistral Small better for coding?

They score almost the same on coding (34.4 vs 34.0); test both on your own repository before choosing.

How many benchmarks do Gemini 1.5 Flash (May 2024) and Mistral Small share?

25 benchmarks have published results for both models. Gemini 1.5 Flash (May 2024) has 42 scored results on Noometry and Mistral Small has 39.

Related comparisons

Go deeper