Meta, proprietary
Muse Spark
Muse Spark by Meta ranks 46th of 354 ranked models on the Noometry Index as of October 2026, with a score of 50.6. Its strongest category is knowledge, where it ranks 13th.
Last verified
Specifications
- Noometry rank
- #46 of 354
- Index score
- 50.6
- Evidence
- Confirmed 27 results
- Provider
Meta
- Released
- April 8, 2026
- Weights
- Proprietary
- Reasoning
- Unknown
- Context window
- —
- Max output
- —
- Input price
- Not listed
- Output price
- Not listed
- Blended price
- Not listed
- Output speed
- Not measured
- Value
- Not ranked
- Knowledge cutoff
- Unknown
Category scores
Each category score combines every public result we have in that category.
- Coding 46.2
- Reasoning 35.9
- Math 47.8
- Knowledge 65.7
- Multimodal 43.4
- Multilingual 56.1
- Instruction Following 75.9
- Long Context 44.4
- Writing & Preference 66.0
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 46.2 | #69 | 2 |
| Reasoning | 35.9 | #67 | 2 |
| Math | 47.8 | #66 | 3 |
| Knowledge | 65.7 | #13 | 3 |
| Multimodal | 43.4 | #24 | 1 |
| Multilingual | 56.1 | #24 | 1 |
| Instruction Following | 75.9 | #51 | 1 |
| Long Context | 44.4 | #69 | 1 |
| Writing & Preference | 66.0 | #39 | 3 |
Strengths and weaknesses
Categories where Muse Spark places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Knowledge | 65.7 | +28.4 | #13 of 314, top 5% |
| Multilingual | 56.1 | +8.7 | #24 of 297, top 9% |
| Writing & Preference | 66.0 | +12.2 | #39 of 312, top 13% |
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Long Context | 44.4 | +3.4 | #69 of 296, top 24% |
| Coding | 46.2 | +7.5 | #69 of 340, top 21% |
| Math | 47.8 | +11.2 | #66 of 327, top 21% |
Closest competitors
The models ranked just above and below Muse Spark. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Qwen3.7 Max | #42 | 51.5 | $3.75 | — | Compare |
| Qwen3.6 Max Preview | #43 | 51.5 | $2.92 | — | Compare |
| GLM-5.2 | #44 | 51.1 | $2.15 | 23 | Compare |
| GPT-5 | #45 | 50.9 | $3.44 | 2 | Compare |
| Claude Opus 4.5 | #47 | 50.5 | $10 | 13 | Compare |
| Muse Spark 1.2 | #48 | 50.3 | $2 | — | Compare |
| MiMo-V2.6-Pro | #49 | 50.3 | $0.54 | — | Compare |
| Claude Sonnet 4.6 | #50 | 50.3 | $6 | — | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| SciCode | 51.5% | #38 of 121, top 32% | Epoch AI | ||
| LMArena Coding | 1481 | #46 of 294, top 16% | LMArena | 2026-10-08 |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| CritPt | 11.3% | #45 of 134, top 34% | Epoch AI | ||
| LMArena Hard Prompts | 1474 | #36 of 297, top 13% | LMArena | 2026-10-08 | |
| Epoch Capabilities Index | 152.04 | #44 of 213, top 21% | Epoch AI | 2026-04-08 |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| OTIS Mock AIME 2024-2025 | 88.9% | #53 of 173, top 31% | Epoch AI | 2026-04-08 | |
| ProofBench | 17% | #53 of 77, top 69% | Epoch AI | ||
| LMArena Math | 1455 | #58 of 285, top 21% | LMArena | 2026-10-08 | |
| FrontierMath (Feb 2025 set) | 39% | #9 of 68, top 14% | Epoch AI | 2026-04-08 | |
| FrontierMath Tier 4 (v1) | 14.6% | #14 of 55, top 26% | Epoch AI | 2026-04-08 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| GPQA Diamond | 89.8% | #41 of 186, top 23% | Epoch AI | 2026-04-08 | |
| Humanity's Last Exam | 40.6% | #6 of 41, top 15% | Epoch AI | ||
| LMArena Expert | 1457 | #64 of 273, top 24% | LMArena | 2026-10-08 |
Multimodal
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Vision | 1306 | #12 of 122, top 10% | LMArena | 2026-10-09 | |
| LMArena Document | 1444 | #24 of 38, top 64% | LMArena | 2026-09-13 |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1464 | #24 of 297, top 9% | LMArena | 2026-10-08 | |
| LMArena Chinese | 1509 | #38 of 285, top 14% | LMArena | 2026-10-08 | |
| LMArena French | 1497 | #14 of 223, top 7% | LMArena | 2026-10-08 | |
| LMArena German | 1497 | #8 of 231, top 4% | LMArena | 2026-10-08 | |
| LMArena Korean | 1459 | #11 of 213, top 6% | LMArena | 2026-10-08 | |
| LMArena Russian | 1466 | #29 of 283, top 11% | LMArena | 2026-10-08 | |
| LMArena Spanish | 1472 | #21 of 226, top 10% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Instruction Following | 1442 | #48 of 298, top 17% | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 1451 | #52 of 291, top 18% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1474 | #21 of 297, top 8% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1459 | #19 of 295, top 7% | LMArena | 2026-10-08 | |
| LMArena Multi-Turn | 1477 | #22 of 295, top 8% | LMArena | 2026-10-08 |
Compare Muse Spark
- Muse Spark vs GPT-5
- Muse Spark vs Claude Opus 4.5
- Muse Spark vs GLM-5.2
- Muse Spark vs Muse Spark 1.2
- Muse Spark vs Qwen3.6 Max Preview
- Muse Spark vs MiMo-V2.6-Pro
- Muse Spark vs GPT-6 Astra
- Muse Spark vs Claude Fable 5.1
- Muse Spark vs Gemini 3.8 Flash
- Muse Spark vs Kimi K3
- Muse Spark vs Grok 4.6
- Muse Spark vs Qwen3.8 Max
- Muse Spark vs GLM-5.3
- Muse Spark vs DeepSeek V4 Pro
Other Meta models
- Muse Spark 1.354.8
- Muse Spark 1.250.3
- Muse Spark 1.149.9
- Muse Glimmer41.7
- Codellama 70b Instruct33.7
- Llama 4 Maverick30.9
- Codellama 34b Instruct30.8
- Llama 3.1-405B30.7
Frequently asked questions
How good is Muse Spark?
Muse Spark by Meta ranks 46th of 354 ranked models on the Noometry Index as of October 2026, with a score of 50.6. Its strongest category is knowledge, where it ranks 13th.
Is Muse Spark open source?
No. Muse Spark is proprietary and available only through Meta's API and partner platforms.
What are Muse Spark's strengths and weaknesses?
Relative to other ranked models, Muse Spark places best in knowledge, multilingual, writing & preference and lowest in long context, coding, math.
What is Muse Spark best at?
Its best category is knowledge, where it ranks 13th on Noometry.