Mistral AI, open weights
Mixtral 8x22B
Mixtral 8x22B by Mistral AI ranks 333rd of 354 ranked models on the Noometry Index as of October 2026, with a score of 27.1. Its strongest category is agentic & tool use, where it ranks 127th. API pricing starts at $2 per million input tokens and $6 per million output tokens, with a 64K-token context window.
Last verified
Specifications
- Noometry rank
- #333 of 354
- Index score
- 27.1
- Evidence
- Confirmed 34 results
- Provider
Mistral AI
- Released
- April 17, 2024
- Weights
- Open weights
- Reasoning
- No
- Context window
- 64K
- Max output
- 64K
- Input price
- $2 / M
- Output price
- $6 / M
- Blended price
- $3 / M
- Output speed
- Not measured
- Value
- #187 of 219
- Knowledge cutoff
- April 2024
- Input
- text
- Hugging Face
- mistralai/Mixtral-8x22B-Instruct-v0.1
Category scores
Each category score combines every public result we have in that category.
- Coding 24.2
- Agentic & Tool Use 23.1
- Reasoning 19.9
- Math 22.9
- Knowledge 15.1
- Multilingual 32.8
- Instruction Following 57.7
- Long Context 34.7
- Writing & Preference 36.9
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 24.2 | #329 | 4 |
| Agentic & Tool Use | 23.1 | #127 | 1 |
| Reasoning | 19.9 | #248 | 2 |
| Math | 22.9 | #275 | 3 |
| Knowledge | 15.1 | #293 | 4 |
| Multilingual | 32.8 | #255 | 1 |
| Instruction Following | 57.7 | #266 | 2 |
| Long Context | 34.7 | #247 | 1 |
| Writing & Preference | 36.9 | #262 | 4 |
Strengths and weaknesses
Categories where Mixtral 8x22B places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Reasoning | 19.9 | −3.7 | #248 of 350, top 71% |
| Agentic & Tool Use | 23.1 | −7.3 | #127 of 154, top 83% |
| Long Context | 34.7 | −6.3 | #247 of 296, top 84% |
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Coding | 24.2 | −14.5 | #329 of 340, top 97% |
| Knowledge | 15.1 | −22.2 | #293 of 314, top 94% |
| Instruction Following | 57.7 | −13.6 | #266 of 305, top 88% |
Closest competitors
The models ranked just above and below Mixtral 8x22B. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Yi-34B | #329 | 27.8 | — | — | Compare |
| Llama 4 Scout | #330 | 27.7 | $0.15 | 272 | Compare |
| Llama 3.2 90B | #331 | 27.5 | — | — | Compare |
| Gemini 1.0 Pro | #332 | 27.3 | — | — | Compare |
| Mixtral 8x7B | #334 | 27.1 | $0.70 | — | Compare |
| Qwen Turbo | #335 | 27.1 | $0.0875 | — | Compare |
| Qwen3-1.7B | #336 | 26.6 | — | — | Compare |
| Mistral Nemo | #337 | 26.4 | $0.15 | — | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| WeirdML | 3.2% | #118 of 119, top 100% | Epoch AI | ||
| BigCodeBench Instruct | 40.6% | #32 of 64, top 50% | BigCodeBench | 2024-04-17 | |
| LMArena Coding | 1166 | #252 of 294, top 86% | LMArena | 2026-10-08 | |
| BigCodeBench Complete | 50.2% | #33 of 66, top 50% | BigCodeBench | 2024-04-17 | |
| BigCodeBench Complete | 45.3% | BigCodeBench | 2024-04-17 | ||
| HumanEval+ | 72% | #18 of 45, top 40% | EvalPlus | ||
| MBPP+ | 64.3% | #20 of 38, top 53% | EvalPlus |
Agentic & Tool Use
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| Cybench | 7.5% | #20 of 21, top 96% | Epoch AI |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Hard Prompts | 1150 | #253 of 297, top 86% | LMArena | 2026-10-08 | |
| DTBench | 55.1% | #122 of 151, top 81% | Epoch AI | ||
| Epoch Capabilities Index | 122.03 | #163 of 213, top 77% | Epoch AI | 2024-04-17 | |
| ForecastBench | 56.3 | #64 of 72, top 89% | Epoch AI |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| Omni-MATH | 16.3% | #52 of 57, top 92% | HELM Capabilities | ||
| LMArena Math | 1184 | #238 of 285, top 84% | LMArena | 2026-10-08 | |
| MATH Level 5 | 24.2% | #59 of 79, top 75% | Epoch AI | 2025-01-27 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| GPQA Diamond | 34.1% | #160 of 186, top 87% | Epoch AI | 2025-01-27 | |
| MMLU-Pro | 46% | #52 of 58, top 90% | HELM Capabilities | ||
| GPQA (HELM) | 33.4% | #50 of 57, top 88% | HELM Capabilities | ||
| LMArena Expert | 1113 | #248 of 273, top 91% | LMArena | 2026-10-08 | |
| MMLU | 77.8% | #27 of 81, top 34% | Epoch AI |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1128 | #255 of 297, top 86% | LMArena | 2026-10-08 | |
| LMArena Chinese | 1116 | #254 of 285, top 90% | LMArena | 2026-10-08 | |
| LMArena French | 1166 | #201 of 223, top 91% | LMArena | 2026-10-08 | |
| LMArena German | 1141 | #204 of 231, top 89% | LMArena | 2026-10-08 | |
| LMArena Japanese | 1037 | #193 of 211, top 92% | LMArena | 2026-10-08 | |
| LMArena Korean | 1057 | #191 of 213, top 90% | LMArena | 2026-10-08 | |
| LMArena Russian | 1158 | #249 of 283, top 88% | LMArena | 2026-10-08 | |
| LMArena Spanish | 1151 | #205 of 226, top 91% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| IFEval | 72.4% | #55 of 57, top 97% | HELM Capabilities | ||
| LMArena Instruction Following | 1147 | #254 of 298, top 86% | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 1144 | #256 of 291, top 88% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1162 | #256 of 297, top 87% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1141 | #254 of 295, top 87% | LMArena | 2026-10-08 | |
| WildBench | 71.1% | #51 of 57, top 90% | HELM Capabilities | ||
| LMArena Multi-Turn | 1130 | #257 of 295, top 88% | LMArena | 2026-10-08 |
API pricing by provider
| Route | Input $/M | Output $/M | Cached input $/M | Checked |
|---|---|---|---|---|
| mistral | $2 | $6 | — | 2026-10-10 |
| openrouter | $2 | $6 | $0.20 | 2026-10-10 |
Compare Mixtral 8x22B
- Mixtral 8x22B vs Mixtral 8x7B
- Mixtral 8x22B vs Gemini 1.0 Pro
- Mixtral 8x22B vs Llama 3.2 90B
- Mixtral 8x22B vs Qwen Turbo
- Mixtral 8x22B vs Llama 4 Scout
- Mixtral 8x22B vs Qwen3-1.7B
- Mixtral 8x22B vs GPT-6 Astra
- Mixtral 8x22B vs Claude Fable 5.1
- Mixtral 8x22B vs Gemini 3.8 Flash
- Mixtral 8x22B vs Kimi K3
- Mixtral 8x22B vs Grok 4.6
- Mixtral 8x22B vs Qwen3.8 Max
- Mixtral 8x22B vs GLM-5.3
- Mixtral 8x22B vs Muse Spark 1.3
Other Mistral AI models
- Mistral Large 443.1
- Mistral Medium 3.540.2
- Mistral Large 339.1
- Mistral Medium36.3
- Magistral Medium35.2
- Devstral Small 250534.3
- Mistral Small33.4
- Pixtral Large32.2
Frequently asked questions
How good is Mixtral 8x22B?
Mixtral 8x22B by Mistral AI ranks 333rd of 354 ranked models on the Noometry Index as of October 2026, with a score of 27.1. Its strongest category is agentic & tool use, where it ranks 127th. API pricing starts at $2 per million input tokens and $6 per million output tokens, with a 64K-token context window.
How much does Mixtral 8x22B cost?
Mixtral 8x22B costs $2 per million input tokens and $6 per million output tokens on Mistral AI's own API.
What is Mixtral 8x22B's context window?
Mixtral 8x22B accepts up to 64K tokens of input and can write up to 64K tokens in one response.
Is Mixtral 8x22B open source?
Yes. Mixtral 8x22B's weights are downloadable from Hugging Face (mistralai/Mixtral-8x22B-Instruct-v0.1); check the license for commercial terms.
What are Mixtral 8x22B's strengths and weaknesses?
Relative to other ranked models, Mixtral 8x22B places best in reasoning, agentic & tool use, long context and lowest in coding, knowledge, instruction following.
What is Mixtral 8x22B best at?
Its best category is agentic & tool use, where it ranks 127th on Noometry.