Mistral AI, unknown
Mistral
Mistral by Mistral AI ranks 303rd of 354 ranked models on the Noometry Index as of October 2026, with a score of 29.9. Its strongest category is reasoning, where it ranks 200th.
Last verified
Specifications
- Noometry rank
- #303 of 354
- Index score
- 29.9
- Evidence
- Confirmed 22 results
- Provider
Mistral AI
- Released
- Unknown
- Weights
- Unknown
- Reasoning
- Unknown
- Context window
- —
- Max output
- —
- Input price
- Not listed
- Output price
- Not listed
- Blended price
- Not listed
- Output speed
- Not measured
- Value
- Not ranked
- Knowledge cutoff
- Unknown
Category scores
Each category score combines every public result we have in that category.
- Coding 33.8
- Reasoning 22.2
- Math 22.3
- Knowledge 16.6
- Multilingual 32.8
- Instruction Following 52.6
- Long Context 35.0
- Writing & Preference 37.0
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 33.8 | #250 | 1 |
| Reasoning | 22.2 | #200 | 1 |
| Math | 22.3 | #278 | 2 |
| Knowledge | 16.6 | #288 | 3 |
| Multilingual | 32.8 | #254 | 1 |
| Instruction Following | 52.6 | #288 | 2 |
| Long Context | 35.0 | #245 | 1 |
| Writing & Preference | 37.0 | #260 | 4 |
Strengths and weaknesses
Categories where Mistral places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Reasoning | 22.2 | −1.4 | #200 of 350, top 58% |
| Coding | 33.8 | −4.9 | #250 of 340, top 74% |
| Long Context | 35.0 | −6.0 | #245 of 296, top 83% |
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Instruction Following | 52.6 | −18.7 | #288 of 305, top 95% |
| Knowledge | 16.6 | −20.8 | #288 of 314, top 92% |
| Multilingual | 32.8 | −14.6 | #254 of 297, top 86% |
Closest competitors
The models ranked just above and below Mistral. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Gemma 7B | #299 | 30.0 | — | — | Compare |
| Qwen2-72B | #300 | 30.0 | — | — | Compare |
| Gemini 1.5 Flash 8B | #301 | 29.9 | — | — | Compare |
| Qwen2.5-VL 72B Instruct | #302 | 29.9 | $4.20 | 43 | Compare |
| OLMo 2 Furious 13B | #304 | 29.7 | — | — | Compare |
| Phi 3 Mini 128k Instruct | #305 | 29.7 | — | — | Compare |
| phi-3-medium 14B | #306 | 29.7 | — | — | Compare |
| Gemma 2B | #307 | 29.6 | — | — | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Coding | 1162 | #254 of 294, top 87% | medium | LMArena | 2026-10-08 |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Hard Prompts | 1149 | #254 of 297, top 86% | medium | LMArena | 2026-10-08 |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| Omni-MATH | 7.2% | #57 of 57, top 100% | HELM Capabilities | ||
| LMArena Math | 1180 | #241 of 285, top 85% | medium | LMArena | 2026-10-08 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| MMLU-Pro | 27.7% | #58 of 58, top 100% | HELM Capabilities | ||
| GPQA (HELM) | 30.3% | #54 of 57, top 95% | HELM Capabilities | ||
| LMArena Expert | 1125 | #246 of 273, top 91% | medium | LMArena | 2026-10-08 |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1129 | #254 of 297, top 86% | medium | LMArena | 2026-10-08 |
| LMArena Chinese | 1109 | #256 of 285, top 90% | medium | LMArena | 2026-10-08 |
| LMArena French | 1180 | #197 of 223, top 89% | medium | LMArena | 2026-10-08 |
| LMArena German | 1155 | #201 of 231, top 88% | medium | LMArena | 2026-10-08 |
| LMArena Japanese | 1013 | #199 of 211, top 95% | medium | LMArena | 2026-10-08 |
| LMArena Korean | 1032 | #196 of 213, top 93% | medium | LMArena | 2026-10-08 |
| LMArena Russian | 1168 | #246 of 283, top 87% | medium | LMArena | 2026-10-08 |
| LMArena Spanish | 1143 | #206 of 226, top 92% | medium | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| IFEval | 56.8% | #57 of 57, top 100% | HELM Capabilities | ||
| LMArena Instruction Following | 1152 | #253 of 298, top 85% | medium | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 1153 | #254 of 291, top 88% | medium | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1165 | #255 of 297, top 86% | medium | LMArena | 2026-10-08 |
| LMArena Creative Writing | 1158 | #247 of 295, top 84% | medium | LMArena | 2026-10-08 |
| WildBench | 66% | #56 of 57, top 99% | HELM Capabilities | ||
| LMArena Multi-Turn | 1147 | #254 of 295, top 87% | medium | LMArena | 2026-10-08 |
Compare Mistral
- Mistral vs Qwen2.5-VL 72B Instruct
- Mistral vs OLMo 2 Furious 13B
- Mistral vs Gemini 1.5 Flash 8B
- Mistral vs Phi 3 Mini 128k Instruct
- Mistral vs Qwen2-72B
- Mistral vs phi-3-medium 14B
- Mistral vs GPT-6 Astra
- Mistral vs Claude Fable 5.1
- Mistral vs Gemini 3.8 Flash
- Mistral vs Kimi K3
- Mistral vs Grok 4.6
- Mistral vs Qwen3.8 Max
- Mistral vs GLM-5.3
- Mistral vs Muse Spark 1.3
Other Mistral AI models
- Mistral Large 443.1
- Mistral Medium 3.540.2
- Mistral Large 339.1
- Mistral Medium36.3
- Magistral Medium35.2
- Devstral Small 250534.3
- Mistral Small33.4
- Pixtral Large32.2
Frequently asked questions
How good is Mistral?
Mistral by Mistral AI ranks 303rd of 354 ranked models on the Noometry Index as of October 2026, with a score of 29.9. Its strongest category is reasoning, where it ranks 200th.
Is Mistral open source?
Mistral AI has not published downloadable weights for Mistral that we can confirm.
What are Mistral's strengths and weaknesses?
Relative to other ranked models, Mistral places best in reasoning, coding, long context and lowest in instruction following, knowledge, multilingual.
What is Mistral best at?
Its best category is reasoning, where it ranks 200th on Noometry.