Meta, proprietary

# Muse Spark 1.1

> Muse Spark 1.1 by Meta, released April 2026. Ranked #51 of 354 with a Noometry Index of 49.9. API: $1.25 in / $4.25 out per M tokens. 1.05M context. Scores, sources and comparisons.
- Canonical page: https://noometry.com/models/muse-spark-1-1
- Last updated: 2026-10-10
- Title: Muse Spark 1.1 Benchmarks, Price & Rank (October 2026)

Muse Spark 1.1 by Meta ranks 51st of 354 ranked models on the Noometry Index as of October 2026, with a score of 49.9. Its strongest category is writing & preference, where it ranks 11th. API pricing starts at $1.25 per million input tokens and $4.25 per million output tokens, with a 1.05M-token context window.

Last verified October 10, 2026

## Specifications

- **Noometry rank:** #51 of 354
- **Index score:** 49.9
- **Evidence:** Confirmed 37 results
- **Provider:** [![](/logos/meta.svg) Meta](https://noometry.com/providers/meta)
- **Released:** April 8, 2026
- **Weights:** Proprietary
- **Reasoning:** Yes
- **Context window:** 1.05M
- **Max output:** 131K
- **Input price:** $1.25 / M
- **Output price:** $4.25 / M
- **Blended price:** $2 / M
- **Output speed:** Not measured
- **Value:** #144 of 219
- **Knowledge cutoff:** Unknown
- **Input:** text, image, pdf, video

## Category scores

Each category score combines every public result we have in that category.

Muse Spark 1.1 category scores

1.  Coding 51.3
2.  Agentic & Tool Use 30.8
3.  Reasoning 47.1
4.  Math 45.5
5.  Knowledge 53.1
6.  Multimodal 42.6
7.  Multilingual 56.7
8.  Instruction Following 76.5
9.  Long Context 44.8
10.  Writing & Preference 73.4
11.  020406080

Muse Spark 1.1 category ranks
| Category | Score | Rank | Results |
| --- | --- | --- | --- |
| [Coding](https://noometry.com/best/coding) | 51.3 | #40 | 4 |
| [Agentic & Tool Use](https://noometry.com/best/agentic) | 30.8 | #73 | 4 |
| [Reasoning](https://noometry.com/best/reasoning) | 47.1 | #44 | 6 |
| [Math](https://noometry.com/best/math) | 45.5 | #76 | 2 |
| [Knowledge](https://noometry.com/best/knowledge) | 53.1 | #59 | 2 |
| [Multimodal](https://noometry.com/best/multimodal) | 42.6 | #29 | 1 |
| [Multilingual](https://noometry.com/best/multilingual) | 56.7 | #17 | 1 |
| [Instruction Following](https://noometry.com/best/instruction-following) | 76.5 | #39 | 1 |
| [Long Context](https://noometry.com/best/long-context) | 44.8 | #58 | 1 |
| [Writing & Preference](https://noometry.com/best/writing) | 73.4 | #11 | 5 |

## Strengths and weaknesses

Categories where Muse Spark 1.1 places highest and lowest among the models ranked in each, with its score against that category's median.

### Strongest categories

Muse Spark 1.1: strongest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Writing & Preference](https://noometry.com/best/writing) | 73.4 | +19.7 | #11 of 312, top 4% |
| [Multilingual](https://noometry.com/best/multilingual) | 56.7 | +9.3 | #17 of 297, top 6% |
| [Coding](https://noometry.com/best/coding) | 51.3 | +12.6 | #40 of 340, top 12% |

### Weakest categories

Muse Spark 1.1: weakest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Agentic & Tool Use](https://noometry.com/best/agentic) | 30.8 | +0.5 | #73 of 154, top 48% |
| [Math](https://noometry.com/best/math) | 45.5 | +8.9 | #76 of 327, top 24% |
| [Multimodal](https://noometry.com/best/multimodal) | 42.6 | +4.1 | #29 of 128, top 23% |

## Closest competitors

The models ranked just above and below Muse Spark 1.1. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Muse Spark 1.1
| Model | Rank | Score | Blended $/M | Speed |  |
| --- | --- | --- | --- | --- | --- |
| [Claude Opus 4.5](https://noometry.com/models/claude-opus-4-5) | #47 | 50.5 | $10 | 13 | [Compare](https://noometry.com/compare/claude-opus-4-5-vs-muse-spark-1-1) |
| [Muse Spark 1.2](https://noometry.com/models/muse-spark-1-2) | #48 | 50.3 | $2 | — | [Compare](https://noometry.com/compare/muse-spark-1-1-vs-muse-spark-1-2) |
| [MiMo-V2.6-Pro](https://noometry.com/models/mimo-v2-6-pro) | #49 | 50.3 | $0.54 | — | [Compare](https://noometry.com/compare/mimo-v2-6-pro-vs-muse-spark-1-1) |
| [Claude Sonnet 4.6](https://noometry.com/models/claude-sonnet-4-6) | #50 | 50.3 | $6 | — | [Compare](https://noometry.com/compare/claude-sonnet-4-6-vs-muse-spark-1-1) |
| [Claude Haiku 5.5](https://noometry.com/models/claude-haiku-5-5) | #52 | 49.5 | $0.20 | — | [Compare](https://noometry.com/compare/claude-haiku-5-5-vs-muse-spark-1-1) |
| [GPT-5.1](https://noometry.com/models/gpt-5-1) | #53 | 49.0 | $3.44 | — | [Compare](https://noometry.com/compare/gpt-5-1-vs-muse-spark-1-1) |
| [Grok 4.20 (Non-Reasoning)](https://noometry.com/models/grok-4-20) | #54 | 48.6 | $1.56 | 61 | [Compare](https://noometry.com/compare/grok-4-20-vs-muse-spark-1-1) |
| [MiMo-V2.6-Flash](https://noometry.com/models/mimo-v2-6-flash) | #55 | 48.5 | $0.18 | — | [Compare](https://noometry.com/compare/mimo-v2-6-flash-vs-muse-spark-1-1) |

Sponsored placements are available on pages like this one. [Advertise on Noometry](https://noometry.com/advertise)

## Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

### Coding

Muse Spark 1.1 Coding benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [DeepSWE](https://noometry.com/benchmarks/deepswe) | 53.3% | #22 of 29, top 76% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena WebDev](https://noometry.com/benchmarks/arena-webdev) | 1542 | #35 of 113, top 31% |  | [LMArena](https://lmarena.ai/leaderboard/webdev) | 2026-10-08 |
| [SciCode](https://noometry.com/benchmarks/scicode) | 58.8% | #13 of 121, top 11% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [SciCode](https://noometry.com/benchmarks/scicode) | 58.8% |  | xhigh | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena Coding](https://noometry.com/benchmarks/arena-coding) | 1498 | #20 of 294, top 7% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Agentic & Tool Use

Muse Spark 1.1 Agentic & Tool Use benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [APEX-Agents](https://noometry.com/benchmarks/apex-agents) | 31.8% | #43 of 49, top 88% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [τ²-bench Banking](https://noometry.com/benchmarks/tau2-banking) | 40.5% | #6 of 26, top 24% | xhigh | [τ²-bench](https://taubench.com/) | 2026-08-04 |
| [GBAEval](https://noometry.com/benchmarks/gbaeval) | 7.9% | #13 of 23, top 57% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [GDP.pdf](https://noometry.com/benchmarks/gdp-pdf) | 15% | #27 of 36, top 75% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [GDP.pdf](https://noometry.com/benchmarks/gdp-pdf) | 15% |  | medium | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [Vending-Bench 2](https://noometry.com/benchmarks/vending-bench-2) | 6,520 | #16 of 60, top 27% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |

### Reasoning

Muse Spark 1.1 Reasoning benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [NYT Connections (extended)](https://noometry.com/benchmarks/nyt-connections) | 84.9% | #30 of 91, top 33% | high reasoning | [Lech Mazur benchmarks](https://github.com/lechmazur/nyt-connections) |  |
| [CritPt](https://noometry.com/benchmarks/critpt) | 15.1% |  |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [CritPt](https://noometry.com/benchmarks/critpt) | 15.1% | #37 of 134, top 28% | xhigh | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena Hard Prompts](https://noometry.com/benchmarks/arena-hard-prompts) | 1486 | #21 of 297, top 8% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [DTBench](https://noometry.com/benchmarks/dtbench) | 94.4% | #23 of 151, top 16% | high | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMCA](https://noometry.com/benchmarks/lmca) | 49.9% | #22 of 125, top 18% | high | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [Surface Evolver Bench](https://noometry.com/benchmarks/surface-evolver-bench) | 52.5% | #15 of 25, top 60% | high | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [Epoch Capabilities Index](https://noometry.com/benchmarks/epoch-capabilities-index) | 154.21 | #37 of 213, top 18% |  | [Epoch AI](https://epoch.ai/eci) | 2026-07-09 |

### Math

Muse Spark 1.1 Math benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [ProofBench](https://noometry.com/benchmarks/proofbench) | 39% | #36 of 77, top 47% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena Math](https://noometry.com/benchmarks/arena-math) | 1483 | #25 of 285, top 9% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Knowledge

Muse Spark 1.1 Knowledge benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [SimpleQA Verified](https://noometry.com/benchmarks/simpleqa-verified) | 57.8% | #17 of 77, top 23% |  | [Epoch AI](https://epoch.ai/benchmarks) | 2026-08-31 |
| [LMArena Expert](https://noometry.com/benchmarks/arena-expert) | 1478 | #43 of 273, top 16% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Multimodal

Muse Spark 1.1 Multimodal benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Vision](https://noometry.com/benchmarks/arena-vision) | 1293 | #21 of 122, top 18% |  | [LMArena](https://lmarena.ai/leaderboard/vision) | 2026-10-09 |
| [LMArena Document](https://noometry.com/benchmarks/arena-document) | 1465 | #15 of 38, top 40% |  | [LMArena](https://lmarena.ai/leaderboard/document) | 2026-09-13 |

### Multilingual

Muse Spark 1.1 Multilingual benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Non-English](https://noometry.com/benchmarks/arena-non-english) | 1472 | #17 of 297, top 6% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Chinese](https://noometry.com/benchmarks/arena-chinese) | 1518 | #29 of 285, top 11% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena French](https://noometry.com/benchmarks/arena-french) | 1494 | #16 of 223, top 8% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena German](https://noometry.com/benchmarks/arena-german) | 1466 | #30 of 231, top 13% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Japanese](https://noometry.com/benchmarks/arena-japanese) | 1451 | #24 of 211, top 12% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Korean](https://noometry.com/benchmarks/arena-korean) | 1458 | #13 of 213, top 7% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Russian](https://noometry.com/benchmarks/arena-russian) | 1483 | #17 of 283, top 7% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Spanish](https://noometry.com/benchmarks/arena-spanish) | 1464 | #31 of 226, top 14% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Instruction Following

Muse Spark 1.1 Instruction Following benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Instruction Following](https://noometry.com/benchmarks/arena-instruction-following) | 1457 | #37 of 298, top 13% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Long Context

Muse Spark 1.1 Long Context benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Longer Query](https://noometry.com/benchmarks/arena-longer-query) | 1462 | #40 of 291, top 14% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Writing & Preference

Muse Spark 1.1 Writing & Preference benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Text](https://noometry.com/benchmarks/arena-text) | 1479 | #18 of 297, top 7% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Creative Writing](https://noometry.com/benchmarks/arena-creative-writing) | 1437 | #42 of 295, top 15% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [EQ-Bench Creative Writing](https://noometry.com/benchmarks/eqbench-creative-writing) | 1927 | #12 of 115, top 11% |  | [EQ-Bench](https://eqbench.com/creative_writing.html) |  |
| [EQ-Bench 4](https://noometry.com/benchmarks/eqbench-4) | 1260 | #8 of 28, top 29% |  | [EQ-Bench](https://eqbench.com/) |  |
| [LMArena Multi-Turn](https://noometry.com/benchmarks/arena-multi-turn) | 1485 | #14 of 295, top 5% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

## API pricing by provider

Muse Spark 1.1 API prices
| Route | Input $/M | Output $/M | Cached input $/M | Checked |
| --- | --- | --- | --- | --- |
| [meta](https://dev.meta.ai/docs) | $1.25 | $4.25 | $0.15 | 2026-10-10 |
| [openrouter](https://openrouter.ai/meta/muse-spark-1.1) | $1.25 | $4.25 | $0.15 | 2026-10-10 |

[All Meta API prices →](https://noometry.com/llm-pricing/meta) [Estimate your cost →](https://noometry.com/tools/cost-calculator)

## Compare Muse Spark 1.1

-   [Muse Spark 1.1 vs Claude Sonnet 4.6](https://noometry.com/compare/claude-sonnet-4-6-vs-muse-spark-1-1)
-   [Muse Spark 1.1 vs Claude Haiku 5.5](https://noometry.com/compare/claude-haiku-5-5-vs-muse-spark-1-1)
-   [Muse Spark 1.1 vs MiMo-V2.6-Pro](https://noometry.com/compare/mimo-v2-6-pro-vs-muse-spark-1-1)
-   [Muse Spark 1.1 vs GPT-5.1](https://noometry.com/compare/gpt-5-1-vs-muse-spark-1-1)
-   [Muse Spark 1.1 vs Muse Spark 1.2](https://noometry.com/compare/muse-spark-1-1-vs-muse-spark-1-2)
-   [Muse Spark 1.1 vs Grok 4.20 (Non-Reasoning)](https://noometry.com/compare/grok-4-20-vs-muse-spark-1-1)
-   [Muse Spark 1.1 vs GPT-6 Astra](https://noometry.com/compare/gpt-6-astra-vs-muse-spark-1-1)
-   [Muse Spark 1.1 vs Claude Fable 5.1](https://noometry.com/compare/claude-fable-5-1-vs-muse-spark-1-1)
-   [Muse Spark 1.1 vs Gemini 3.8 Flash](https://noometry.com/compare/gemini-3-8-flash-vs-muse-spark-1-1)
-   [Muse Spark 1.1 vs Kimi K3](https://noometry.com/compare/kimi-k3-vs-muse-spark-1-1)
-   [Muse Spark 1.1 vs Grok 4.6](https://noometry.com/compare/grok-4-6-vs-muse-spark-1-1)
-   [Muse Spark 1.1 vs Qwen3.8 Max](https://noometry.com/compare/muse-spark-1-1-vs-qwen3-8-max)
-   [Muse Spark 1.1 vs GLM-5.3](https://noometry.com/compare/glm-5-3-vs-muse-spark-1-1)
-   [Muse Spark 1.1 vs DeepSeek V4 Pro](https://noometry.com/compare/deepseek-v4-pro-vs-muse-spark-1-1)

## Other Meta models

-   [Muse Spark 1.3](https://noometry.com/models/muse-spark-1-3)54.8
-   [Muse Spark](https://noometry.com/models/muse-spark)50.6
-   [Muse Spark 1.2](https://noometry.com/models/muse-spark-1-2)50.3
-   [Muse Glimmer](https://noometry.com/models/muse-glimmer)41.7
-   [Codellama 70b Instruct](https://noometry.com/models/codellama-70b-instruct)33.7
-   [Llama 4 Maverick](https://noometry.com/models/llama-4-maverick)30.9
-   [Codellama 34b Instruct](https://noometry.com/models/codellama-34b-instruct)30.8
-   [Llama 3.1-405B](https://noometry.com/models/llama-3-1-405b)30.7

## Frequently asked questions

### How good is Muse Spark 1.1?

Muse Spark 1.1 by Meta ranks 51st of 354 ranked models on the Noometry Index as of October 2026, with a score of 49.9. Its strongest category is writing & preference, where it ranks 11th. API pricing starts at $1.25 per million input tokens and $4.25 per million output tokens, with a 1.05M-token context window.

### How much does Muse Spark 1.1 cost?

Muse Spark 1.1 costs $1.25 per million input tokens and $4.25 per million output tokens on Meta's own API, with cached input at $0.15.

### What is Muse Spark 1.1's context window?

Muse Spark 1.1 accepts up to 1.05M tokens of input and can write up to 131K tokens in one response.

### Is Muse Spark 1.1 open source?

No. Muse Spark 1.1 is proprietary and available only through Meta's API and partner platforms.

### What are Muse Spark 1.1's strengths and weaknesses?

Relative to other ranked models, Muse Spark 1.1 places best in writing & preference, multilingual, coding and lowest in agentic & tool use, math, multimodal.

### What is Muse Spark 1.1 best at?

Its best category is writing & preference, where it ranks 11th on Noometry.

### Cite this page

Noometry. (2026). Muse Spark 1.1 benchmarks and pricing. Retrieved October 10, 2026, from https://noometry.com/models/muse-spark-1-1

Quote Noometry with a link back to this page. It is also available in [Markdown](https://noometry.com/md/models/muse-spark-1-1.md).
