Model comparison

Muse Spark 1.3 vs o1-pro

Muse Spark 1.3 is the stronger model overall, scoring 54.8 to 31.5 on the Noometry Index.

Last verified . 0 shared benchmarks.

Muse Spark 1.3 Meta

54.8

Rank #27 Confirmed

o1-pro OpenAI

31.5

Rank #271 Reported

Summary

  • The widest gap is in reasoning, where Muse Spark 1.3 leads 54.0 to 20.4.
  • Muse Spark 1.3 is cheaper at $1.25 / $4.25 per million input/output tokens, against $150 / $600 for o1-pro.
  • Muse Spark 1.3 accepts more context: 1.05M tokens versus 200K.

Side by side

Muse Spark 1.3 and o1-pro specifications
Muse Spark 1.3o1-pro
ProviderMetaOpenAI
Noometry Index54.831.5
Released2026-09-022025-03-19
WeightsProprietaryProprietary
Context window1.05M200K
Max output131K100K
Input $ / M tokens$1.25$150
Output $ / M tokens$4.25$600
Results tracked373

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Muse Spark 1.3: 56.6 (#21), o1-pro: —

Coding benchmarks
BenchmarkMuse Spark 1.3o1-pro
CursorBench41.6%—
LMArena WebDev1657—
SciCode59.7%—
LMArena Coding1514—

Agentic & Tool Use Not comparable

Muse Spark 1.3: 38.6 (#30), o1-pro: —

Agentic & Tool Use benchmarks
BenchmarkMuse Spark 1.3o1-pro
APEX-Agents57.8%—
GDP.pdf27.6%—

Reasoning Muse Spark 1.3 leads

Muse Spark 1.3: 54.0 (#27), o1-pro: 20.4 (#239)

Reasoning benchmarks
BenchmarkMuse Spark 1.3o1-pro
NYT Connections (extended)85.1%—
ARC-AGI-1—23.3%
CritPt26%—
Chess Puzzles38%—
EnigmaEval—6.1%
LMArena Hard Prompts1503—
Mystery Game Puzzles25%—
DTBench96.5%—
LMCA53.9%—
Bench to the Future 30.14—
Epoch Capabilities Index156.75—

Math Not comparable

Muse Spark 1.3: 73.1 (#21), o1-pro: —

Math benchmarks
BenchmarkMuse Spark 1.3o1-pro
FrontierMath (Tiers 1-3)74.4%—
FrontierMath Tier 446.3%—
OTIS Mock AIME 2024-202599.2%—
ProofBench58%—
LMArena Math1494—

Knowledge Muse Spark 1.3 leads

Muse Spark 1.3: 42.6 (#95), o1-pro: 29.7 (#234)

Knowledge benchmarks
BenchmarkMuse Spark 1.3o1-pro
Humanity's Last Exam—8.1%
LMArena Expert1516—

Multimodal Not comparable

Muse Spark 1.3: 43.7 (#22), o1-pro: —

Multimodal benchmarks
BenchmarkMuse Spark 1.3o1-pro
LMArena Vision1309—
LMArena Document1471—

Multilingual Not comparable

Muse Spark 1.3: 57.4 (#8), o1-pro: —

Multilingual benchmarks
BenchmarkMuse Spark 1.3o1-pro
LMArena Non-English1481—
LMArena Chinese1529—
LMArena French1524—
LMArena German1515—
LMArena Japanese1474—
LMArena Korean1501—
LMArena Russian1490—
LMArena Spanish1490—

Instruction Following Not comparable

Muse Spark 1.3: 77.5 (#22), o1-pro: —

Instruction Following benchmarks
BenchmarkMuse Spark 1.3o1-pro
LMArena Instruction Following1477—

Long Context Not comparable

Muse Spark 1.3: 45.6 (#32), o1-pro: —

Long Context benchmarks
BenchmarkMuse Spark 1.3o1-pro
LMArena Longer Query1488—

Writing & Preference Not comparable

Muse Spark 1.3: 73.6 (#9), o1-pro: —

Writing & Preference benchmarks
BenchmarkMuse Spark 1.3o1-pro
LMArena Text1490—
LMArena Creative Writing1455—
EQ-Bench Creative Writing1906—
LMArena Multi-Turn1482—

Frequently asked questions

Is Muse Spark 1.3 better than o1-pro?

Muse Spark 1.3 is the stronger model overall, scoring 54.8 to 31.5 on the Noometry Index.

Which is cheaper, Muse Spark 1.3 or o1-pro?

Muse Spark 1.3 is cheaper. It lists at $1.25 per million input tokens and $4.25 per million output tokens; o1-pro lists at $150 and $600.

Which has the bigger context window?

Muse Spark 1.3 does, with 1.05M tokens against 200K.

How many benchmarks do Muse Spark 1.3 and o1-pro share?

0 benchmarks have published results for both models. Muse Spark 1.3 has 37 scored results on Noometry and o1-pro has 3.

Related comparisons

Go deeper