Model comparison

MiMo-V2.6-Pro vs Muse Spark 1.1

MiMo-V2.6-Pro and Muse Spark 1.1 score almost the same on the Noometry Index (50.3 vs 49.9), so choose on price, context window or the category you care about most.

Last verified . 18 shared benchmarks.

MiMo-V2.6-Pro Xiaomi

50.3

Rank #49 Confirmed

Muse Spark 1.1 Meta

49.9

Rank #51 Confirmed

Summary

  • They share 18 benchmarks with published results for both. MiMo-V2.6-Pro scores higher in 6 categories and Muse Spark 1.1 in 4 categories; 9 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Muse Spark 1.1 leads 53.1 to 43.5.
  • The biggest single-benchmark swing is ProofBench: 70% for MiMo-V2.6-Pro and 39% for Muse Spark 1.1.
  • MiMo-V2.6-Pro is cheaper at $0.43 / $0.87 per million input/output tokens, against $1.25 / $4.25 for Muse Spark 1.1.
  • MiMo-V2.6-Pro has downloadable open weights; the other is API-only.

Side by side

MiMo-V2.6-Pro and Muse Spark 1.1 specifications
MiMo-V2.6-ProMuse Spark 1.1
ProviderXiaomiMeta
Noometry Index50.349.9
Released2026-09-212026-04-08
WeightsOpenProprietary
Context window1.05M1.05M
Max output131K131K
Input $ / M tokens$0.43$1.25
Output $ / M tokens$0.87$4.25
Results tracked1937

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding MiMo-V2.6-Pro leads

MiMo-V2.6-Pro: 55.5 (#23), Muse Spark 1.1: 51.3 (#40)

Coding benchmarks
BenchmarkMiMo-V2.6-ProMuse Spark 1.1
LMArena WebDev16291542
SciCode60.9%58.8%
LMArena Coding15341498
DeepSWE—53.3%
ALE-Bench1,158—

Agentic & Tool Use MiMo-V2.6-Pro leads

MiMo-V2.6-Pro: 37.5 (#35), Muse Spark 1.1: 30.8 (#73)

Agentic & Tool Use benchmarks
BenchmarkMiMo-V2.6-ProMuse Spark 1.1
APEX-Agents59.5%31.8%
τ²-bench Banking—40.5%
GBAEval—7.9%
GDP.pdf—15%
Vending-Bench 2—6,520

Reasoning Muse Spark 1.1 leads

MiMo-V2.6-Pro: 43.1 (#50), Muse Spark 1.1: 47.1 (#44)

Reasoning benchmarks
BenchmarkMiMo-V2.6-ProMuse Spark 1.1
CritPt26.6%15.1%
LMArena Hard Prompts15121486
NYT Connections (extended)—84.9%
DTBench—94.4%
LMCA—49.9%
Surface Evolver Bench—52.5%
Epoch Capabilities Index—154.21

Math MiMo-V2.6-Pro leads

MiMo-V2.6-Pro: 54.5 (#45), Muse Spark 1.1: 45.5 (#76)

Math benchmarks
BenchmarkMiMo-V2.6-ProMuse Spark 1.1
ProofBench70%39%
LMArena Math14941483

Knowledge Muse Spark 1.1 leads

MiMo-V2.6-Pro: 43.5 (#92), Muse Spark 1.1: 53.1 (#59)

Knowledge benchmarks
BenchmarkMiMo-V2.6-ProMuse Spark 1.1
LMArena Expert15431478
SimpleQA Verified—57.8%

Multimodal Muse Spark 1.1 leads

MiMo-V2.6-Pro: 40.8 (#43), Muse Spark 1.1: 42.6 (#29)

Multimodal benchmarks
BenchmarkMiMo-V2.6-ProMuse Spark 1.1
LMArena Vision12641293
LMArena Document—1465

Multilingual Too close to call

MiMo-V2.6-Pro: 56.9 (#14), Muse Spark 1.1: 56.7 (#17)

Multilingual benchmarks
BenchmarkMiMo-V2.6-ProMuse Spark 1.1
LMArena Non-English14741472
LMArena Chinese15291518
LMArena Russian14801483
LMArena French—1494
LMArena German—1466
LMArena Japanese—1451
LMArena Korean—1458
LMArena Spanish—1464

Instruction Following MiMo-V2.6-Pro leads

MiMo-V2.6-Pro: 78.2 (#12), Muse Spark 1.1: 76.5 (#39)

Instruction Following benchmarks
BenchmarkMiMo-V2.6-ProMuse Spark 1.1
LMArena Instruction Following14931457

Long Context MiMo-V2.6-Pro leads

MiMo-V2.6-Pro: 46.0 (#27), Muse Spark 1.1: 44.8 (#58)

Long Context benchmarks
BenchmarkMiMo-V2.6-ProMuse Spark 1.1
LMArena Longer Query15011462

Writing & Preference Muse Spark 1.1 leads

MiMo-V2.6-Pro: 66.8 (#33), Muse Spark 1.1: 73.4 (#11)

Writing & Preference benchmarks
BenchmarkMiMo-V2.6-ProMuse Spark 1.1
LMArena Text14921479
LMArena Creative Writing14681437
LMArena Multi-Turn14641485
EQ-Bench Creative Writing—1927
EQ-Bench 4—1260

Frequently asked questions

Is MiMo-V2.6-Pro better than Muse Spark 1.1?

MiMo-V2.6-Pro and Muse Spark 1.1 score almost the same on the Noometry Index (50.3 vs 49.9), so choose on price, context window or the category you care about most.

Which is cheaper, MiMo-V2.6-Pro or Muse Spark 1.1?

MiMo-V2.6-Pro is cheaper. It lists at $0.43 per million input tokens and $0.87 per million output tokens; Muse Spark 1.1 lists at $1.25 and $4.25.

Is MiMo-V2.6-Pro or Muse Spark 1.1 better for coding?

MiMo-V2.6-Pro scores higher on coding benchmarks: 55.5 versus 51.3 in the Noometry coding category.

Which has the bigger context window?

Both accept 1.05M tokens.

How many benchmarks do MiMo-V2.6-Pro and Muse Spark 1.1 share?

18 benchmarks have published results for both models. MiMo-V2.6-Pro has 19 scored results on Noometry and Muse Spark 1.1 has 37.

Related comparisons

Go deeper