Model comparison

Ministral 8B vs Qwen1.5 4b Chat

Ministral 8B and Qwen1.5 4b Chat score almost the same on the Noometry Index (28.2 vs 28.8), so choose on price, context window or the category you care about most.

Last verified . 12 shared benchmarks.

Ministral 8B Mistral AI

28.2

Rank #325 Confirmed

Qwen1.5 4b Chat Alibaba (Qwen)

28.8

Rank #322 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Ministral 8B scores higher in 5 categories and Qwen1.5 4b Chat in 3 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Ministral 8B leads 39.6 to 23.8.

Side by side

Ministral 8B and Qwen1.5 4b Chat specifications
Ministral 8BQwen1.5 4b Chat
ProviderMistral AIAlibaba (Qwen)
Noometry Index28.228.8
Released2024-10-01—
WeightsOpenOpen
Context window262K—
Max output262K—
Input $ / M tokens$0.15—
Output $ / M tokens$0.15—
Results tracked1713

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Ministral 8B leads

Ministral 8B: 35.0 (#230), Qwen1.5 4b Chat: 29.1 (#308)

Coding benchmarks
BenchmarkMinistral 8BQwen1.5 4b Chat
LMArena Coding1202999

Agentic & Tool Use Not comparable

Ministral 8B: 16.4 (#148), Qwen1.5 4b Chat: —

Agentic & Tool Use benchmarks
BenchmarkMinistral 8BQwen1.5 4b Chat
Berkeley Function Calling Leaderboard11.1%—

Reasoning Too close to call

Ministral 8B: 18.4 (#281), Qwen1.5 4b Chat: 18.5 (#279)

Reasoning benchmarks
BenchmarkMinistral 8BQwen1.5 4b Chat
LMArena Hard Prompts1191976
DTBench45.7%—

Math Qwen1.5 4b Chat leads

Ministral 8B: 25.7 (#267), Qwen1.5 4b Chat: 30.4 (#234)

Math benchmarks
BenchmarkMinistral 8BQwen1.5 4b Chat
LMArena Math11881026
MATH Level 514.9%—

Knowledge Qwen1.5 4b Chat leads

Ministral 8B: 12.6 (#297), Qwen1.5 4b Chat: 26.7 (#255)

Knowledge benchmarks
BenchmarkMinistral 8BQwen1.5 4b Chat
LMArena Expert1170980
GPQA Diamond27.1%—
Vectara Hallucination Rate7.4%—

Multilingual Ministral 8B leads

Ministral 8B: 35.1 (#247), Qwen1.5 4b Chat: 24.1 (#290)

Multilingual benchmarks
BenchmarkMinistral 8BQwen1.5 4b Chat
LMArena Non-English1165979
LMArena Chinese11931024
LMArena Russian1195952
LMArena German—902

Instruction Following Ministral 8B leads

Ministral 8B: 60.5 (#250), Qwen1.5 4b Chat: 49.0 (#300)

Instruction Following benchmarks
BenchmarkMinistral 8BQwen1.5 4b Chat
LMArena Instruction Following1161978

Long Context Ministral 8B leads

Ministral 8B: 36.7 (#227), Qwen1.5 4b Chat: 30.1 (#290)

Long Context benchmarks
BenchmarkMinistral 8BQwen1.5 4b Chat
LMArena Longer Query1212988

Writing & Preference Ministral 8B leads

Ministral 8B: 39.6 (#246), Qwen1.5 4b Chat: 23.8 (#309)

Writing & Preference benchmarks
BenchmarkMinistral 8BQwen1.5 4b Chat
LMArena Text1191997
LMArena Creative Writing1175969
LMArena Multi-Turn1166977

Frequently asked questions

Is Ministral 8B better than Qwen1.5 4b Chat?

Ministral 8B and Qwen1.5 4b Chat score almost the same on the Noometry Index (28.2 vs 28.8), so choose on price, context window or the category you care about most.

Is Ministral 8B or Qwen1.5 4b Chat better for coding?

Ministral 8B scores higher on coding benchmarks: 35.0 versus 29.1 in the Noometry coding category.

How many benchmarks do Ministral 8B and Qwen1.5 4b Chat share?

12 benchmarks have published results for both models. Ministral 8B has 17 scored results on Noometry and Qwen1.5 4b Chat has 13.

Related comparisons

Go deeper