Mistral AI, open weights
Mistral 7B
Mistral 7B by Mistral AI ranks 351st of 354 ranked models on the Noometry Index as of October 2026, with a score of 23.0. Its strongest category is long context, where it ranks 271st. API pricing starts at $0.25 per million input tokens and $0.25 per million output tokens, with a 8K-token context window.
Last verified
Specifications
- Noometry rank
- #351 of 354
- Index score
- 23.0
- Evidence
- Confirmed 37 results
- Provider
Mistral AI
- Released
- September 27, 2023
- Weights
- Open weights
- Reasoning
- No
- Context window
- 8K
- Max output
- 8K
- Input price
- $0.25 / M
- Output price
- $0.25 / M
- Blended price
- $0.25 / M
- Output speed
- Not measured
- Value
- #67 of 219
- Knowledge cutoff
- December 2023
- Input
- text
Category scores
Each category score combines every public result we have in that category.
- Coding 26.4
- Reasoning 13.1
- Math 8.1
- Knowledge 7.4
- Multilingual 25.8
- Instruction Following 54.2
- Long Context 32.2
- Writing & Preference 30.7
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 26.4 | #326 | 3 |
| Reasoning | 13.1 | #336 | 3 |
| Math | 8.1 | #325 | 3 |
| Knowledge | 7.4 | #311 | 2 |
| Multilingual | 25.8 | #283 | 1 |
| Instruction Following | 54.2 | #280 | 1 |
| Long Context | 32.2 | #271 | 1 |
| Writing & Preference | 30.7 | #286 | 3 |
Strengths and weaknesses
Categories where Mistral 7B places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Long Context | 32.2 | −8.7 | #271 of 296, top 92% |
| Writing & Preference | 30.7 | −23.1 | #286 of 312, top 92% |
| Instruction Following | 54.2 | −17.1 | #280 of 305, top 92% |
Closest competitors
The models ranked just above and below Mistral 7B. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Claude 2 | #346 | 25.0 | — | — | Compare |
| DeepSeek LLM 67B | #347 | 24.9 | — | — | Compare |
| Llama 13b | #348 | 24.4 | — | — | Compare |
| Llama 2-70B | #349 | 24.4 | — | — | Compare |
| GPT-3.5-turbo | #350 | 23.2 | $0.75 | — | Compare |
| Llama 3.1-8B | #352 | 23.0 | $0.0575 | — | Compare |
| Gemma 3 1B | #353 | 21.1 | — | — | Compare |
| Llama 3.2 1B | #354 | 20.1 | $0.0705 | — | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| BigCodeBench Instruct | 19.5% | #62 of 64, top 97% | BigCodeBench | 2024-05-22 | |
| LMArena Coding | 1082 | #276 of 294, top 94% | LMArena | 2026-10-08 | |
| LMArena Coding | 1018 | LMArena | 2026-10-08 | ||
| BigCodeBench Complete | 23.5% | BigCodeBench | 2024-05-22 | ||
| BigCodeBench Complete | 27.3% | #63 of 66, top 96% | BigCodeBench | 2024-05-22 | |
| HumanEval+ | 36% | #39 of 45, top 87% | EvalPlus | ||
| HumanEval+ | 23.8% | EvalPlus | |||
| MBPP+ | 37% | EvalPlus | |||
| MBPP+ | 42.1% | #36 of 38, top 95% | EvalPlus |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| Chess Puzzles | 0% | #123 of 129, top 96% | Epoch AI | 2026-08-30 | |
| LMArena Hard Prompts | 1002 | LMArena | 2026-10-08 | ||
| LMArena Hard Prompts | 1067 | #278 of 297, top 94% | LMArena | 2026-10-08 | |
| DTBench | 42.5% | #149 of 151, top 99% | Epoch AI | ||
| Adversarial NLI | 47.1% | #8 of 9, top 89% | Epoch AI | ||
| BIG-Bench Hard | 56.1% | #16 of 27, top 60% | Epoch AI | ||
| Epoch Capabilities Index | 112.21 | #188 of 213, top 89% | Epoch AI | 2023-09-27 | |
| Epoch Capabilities Index | 108.99 | Epoch AI | 2024-05-27 | ||
| HellaSwag | 81% | #16 of 29, top 56% | Epoch AI | ||
| PIQA | 83% | #12 of 27, top 45% | Epoch AI | ||
| PIQA | 82.2% | Epoch AI | |||
| PIQA | 82.2% | Epoch AI | |||
| WinoGrande | 75.3% | #23 of 43, top 54% | Epoch AI |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| OTIS Mock AIME 2024-2025 | 0.3% | #172 of 173, top 100% | Epoch AI | 2026-08-30 | |
| LMArena Math | 1027 | LMArena | 2026-10-08 | ||
| LMArena Math | 1085 | #270 of 285, top 95% | LMArena | 2026-10-08 | |
| MATH Level 5 | 3.7% | #78 of 79, top 99% | Epoch AI | 2025-01-27 | |
| MATH Level 5 | 3.6% | Epoch AI | 2025-01-27 | ||
| GSM8K | 54.4% | #20 of 38, top 53% | Epoch AI | ||
| GSM8K | 35.4% | Epoch AI |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| GPQA Diamond | 15.2% | #185 of 186, top 100% | Epoch AI | 2025-01-27 | |
| GPQA Diamond | 13.2% | Epoch AI | 2025-01-27 | ||
| LMArena Expert | 954 | LMArena | 2026-10-08 | ||
| LMArena Expert | 1036 | #267 of 273, top 98% | LMArena | 2026-10-08 | |
| ARC (AI2) Challenge | 78.6% | #13 of 39, top 34% | Epoch AI | ||
| BoolQ | 83.2% | Epoch AI | |||
| BoolQ | 87.4% | #4 of 23, top 18% | Epoch AI | ||
| MMLU | 62.5% | #60 of 81, top 75% | Epoch AI | ||
| MMLU | 62.5% | #60 of 81, top 75% | Epoch AI | ||
| MMLU | 59.9% | Epoch AI | |||
| OpenBookQA | 79.8% | #7 of 19, top 37% | Epoch AI | ||
| TriviaQA | 75.2% | #15 of 25, top 60% | Epoch AI |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 944 | LMArena | 2026-10-08 | ||
| LMArena Non-English | 1012 | #283 of 297, top 96% | LMArena | 2026-10-08 | |
| LMArena Chinese | 1009 | #277 of 285, top 98% | LMArena | 2026-10-08 | |
| LMArena Chinese | 932 | LMArena | 2026-10-08 | ||
| LMArena French | 946 | LMArena | 2026-10-08 | ||
| LMArena French | 1037 | #221 of 223, top 100% | LMArena | 2026-10-08 | |
| LMArena German | 987 | #228 of 231, top 99% | LMArena | 2026-10-08 | |
| LMArena German | 931 | LMArena | 2026-10-08 | ||
| LMArena Japanese | 878 | #211 of 211, top 100% | LMArena | 2026-10-08 | |
| LMArena Russian | 1001 | LMArena | 2026-10-08 | ||
| LMArena Russian | 1018 | #273 of 283, top 97% | LMArena | 2026-10-08 | |
| LMArena Spanish | 1026 | #225 of 226, top 100% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Instruction Following | 1003 | LMArena | 2026-10-08 | ||
| LMArena Instruction Following | 1060 | #276 of 298, top 93% | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 1060 | #276 of 291, top 95% | LMArena | 2026-10-08 | |
| LMArena Longer Query | 1001 | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1024 | LMArena | 2026-10-08 | ||
| LMArena Text | 1090 | #276 of 297, top 93% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1028 | LMArena | 2026-10-08 | ||
| LMArena Creative Writing | 1068 | #275 of 295, top 94% | LMArena | 2026-10-08 | |
| LMArena Multi-Turn | 1008 | LMArena | 2026-10-08 | ||
| LMArena Multi-Turn | 1062 | #273 of 295, top 93% | LMArena | 2026-10-08 |
API pricing by provider
| Route | Input $/M | Output $/M | Cached input $/M | Checked |
|---|---|---|---|---|
| mistral | $0.25 | $0.25 | — | 2026-10-10 |
Compare Mistral 7B
- Mistral 7B vs GPT-3.5-turbo
- Mistral 7B vs Llama 3.1-8B
- Mistral 7B vs Llama 2-70B
- Mistral 7B vs Gemma 3 1B
- Mistral 7B vs Llama 13b
- Mistral 7B vs Llama 3.2 1B
- Mistral 7B vs GPT-6 Astra
- Mistral 7B vs Claude Fable 5.1
- Mistral 7B vs Gemini 3.8 Flash
- Mistral 7B vs Kimi K3
- Mistral 7B vs Grok 4.6
- Mistral 7B vs Qwen3.8 Max
- Mistral 7B vs GLM-5.3
- Mistral 7B vs Muse Spark 1.3
Other Mistral AI models
- Mistral Large 443.1
- Mistral Medium 3.540.2
- Mistral Large 339.1
- Mistral Medium36.3
- Magistral Medium35.2
- Devstral Small 250534.3
- Mistral Small33.4
- Pixtral Large32.2
Frequently asked questions
How good is Mistral 7B?
Mistral 7B by Mistral AI ranks 351st of 354 ranked models on the Noometry Index as of October 2026, with a score of 23.0. Its strongest category is long context, where it ranks 271st. API pricing starts at $0.25 per million input tokens and $0.25 per million output tokens, with a 8K-token context window.
How much does Mistral 7B cost?
Mistral 7B costs $0.25 per million input tokens and $0.25 per million output tokens on Mistral AI's own API.
What is Mistral 7B's context window?
Mistral 7B accepts up to 8K tokens of input and can write up to 8K tokens in one response.
Is Mistral 7B open source?
Yes. Mistral 7B's weights are downloadable; check the license for commercial terms.
What are Mistral 7B's strengths and weaknesses?
Relative to other ranked models, Mistral 7B places best in long context, writing & preference, instruction following and lowest in math, knowledge, reasoning.
What is Mistral 7B best at?
Its best category is long context, where it ranks 271st on Noometry.