Model comparison

Granite 4.0 Micro vs Muse Spark 1.3

Muse Spark 1.3 is the stronger model overall, scoring 54.8 to 29.0 on the Noometry Index. Granite 4.0 Micro costs 49× less per token, which makes it the better buy when Muse Spark 1.3's lead doesn't matter for your workload.

Last verified . 2 shared benchmarks.

Granite 4.0 Micro IBM

29.0

Rank #318 Confirmed

Muse Spark 1.3 Meta

54.8

Rank #27 Confirmed

Summary

  • They share 2 benchmarks with published results for both. Granite 4.0 Micro scores higher in 0 categories and Muse Spark 1.3 in 5 categories; 5 gaps are clear of the uncertainty.
  • The widest gap is in math, where Muse Spark 1.3 leads 73.1 to 12.0.
  • The biggest single-benchmark swing is OTIS Mock AIME 2024-2025: 2.8% for Granite 4.0 Micro and 99.2% for Muse Spark 1.3.
  • Granite 4.0 Micro is cheaper at $0.017 / $0.11 per million input/output tokens, against $1.25 / $4.25 for Muse Spark 1.3.
  • Muse Spark 1.3 accepts more context: 1.05M tokens versus 131K.
  • Granite 4.0 Micro has downloadable open weights; the other is API-only.

Side by side

Granite 4.0 Micro and Muse Spark 1.3 specifications
Granite 4.0 MicroMuse Spark 1.3
ProviderIBMMeta
Noometry Index29.054.8
Released2025-10-022026-09-02
WeightsOpenProprietary
Context window131K1.05M
Max output118K131K
Input $ / M tokens$0.017$1.25
Output $ / M tokens$0.11$4.25
Results tracked837

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Granite 4.0 Micro: —, Muse Spark 1.3: 56.6 (#21)

Coding benchmarks
BenchmarkGranite 4.0 MicroMuse Spark 1.3
CursorBench—41.6%
LMArena WebDev—1657
SciCode—59.7%
LMArena Coding—1514

Agentic & Tool Use Not comparable

Granite 4.0 Micro: —, Muse Spark 1.3: 38.6 (#30)

Agentic & Tool Use benchmarks
BenchmarkGranite 4.0 MicroMuse Spark 1.3
APEX-Agents—57.8%
GDP.pdf—27.6%

Reasoning Muse Spark 1.3 leads

Granite 4.0 Micro: 19.2 (#265), Muse Spark 1.3: 54.0 (#27)

Reasoning benchmarks
BenchmarkGranite 4.0 MicroMuse Spark 1.3
Chess Puzzles0%38%
NYT Connections (extended)—85.1%
CritPt—26%
LMArena Hard Prompts—1503
Mystery Game Puzzles—25%
DTBench—96.5%
LMCA—53.9%
Bench to the Future 3—0.14
Epoch Capabilities Index—156.75

Math Muse Spark 1.3 leads

Granite 4.0 Micro: 12.0 (#307), Muse Spark 1.3: 73.1 (#21)

Math benchmarks
BenchmarkGranite 4.0 MicroMuse Spark 1.3
OTIS Mock AIME 2024-20252.8%99.2%
FrontierMath (Tiers 1-3)—74.4%
FrontierMath Tier 4—46.3%
ProofBench—58%
Omni-MATH20.9%—
LMArena Math—1494

Knowledge Muse Spark 1.3 leads

Granite 4.0 Micro: 9.9 (#304), Muse Spark 1.3: 42.6 (#95)

Knowledge benchmarks
BenchmarkGranite 4.0 MicroMuse Spark 1.3
GPQA Diamond28.3%—
MMLU-Pro39.5%—
GPQA (HELM)30.7%—
LMArena Expert—1516

Multimodal Not comparable

Granite 4.0 Micro: —, Muse Spark 1.3: 43.7 (#22)

Multimodal benchmarks
BenchmarkGranite 4.0 MicroMuse Spark 1.3
LMArena Vision—1309
LMArena Document—1471

Multilingual Not comparable

Granite 4.0 Micro: —, Muse Spark 1.3: 57.4 (#8)

Multilingual benchmarks
BenchmarkGranite 4.0 MicroMuse Spark 1.3
LMArena Non-English—1481
LMArena Chinese—1529
LMArena French—1524
LMArena German—1515
LMArena Japanese—1474
LMArena Korean—1501
LMArena Russian—1490
LMArena Spanish—1490

Instruction Following Muse Spark 1.3 leads

Granite 4.0 Micro: 69.9 (#169), Muse Spark 1.3: 77.5 (#22)

Instruction Following benchmarks
BenchmarkGranite 4.0 MicroMuse Spark 1.3
IFEval84.9%—
LMArena Instruction Following—1477

Long Context Not comparable

Granite 4.0 Micro: —, Muse Spark 1.3: 45.6 (#32)

Long Context benchmarks
BenchmarkGranite 4.0 MicroMuse Spark 1.3
LMArena Longer Query—1488

Writing & Preference Muse Spark 1.3 leads

Granite 4.0 Micro: 46.7 (#216), Muse Spark 1.3: 73.6 (#9)

Writing & Preference benchmarks
BenchmarkGranite 4.0 MicroMuse Spark 1.3
LMArena Text—1490
LMArena Creative Writing—1455
EQ-Bench Creative Writing—1906
WildBench67%—
LMArena Multi-Turn—1482

Frequently asked questions

Is Granite 4.0 Micro better than Muse Spark 1.3?

Muse Spark 1.3 is the stronger model overall, scoring 54.8 to 29.0 on the Noometry Index. Granite 4.0 Micro costs 49× less per token, which makes it the better buy when Muse Spark 1.3's lead doesn't matter for your workload.

Which is cheaper, Granite 4.0 Micro or Muse Spark 1.3?

Granite 4.0 Micro is cheaper. It lists at $0.017 per million input tokens and $0.11 per million output tokens; Muse Spark 1.3 lists at $1.25 and $4.25.

Which has the bigger context window?

Muse Spark 1.3 does, with 1.05M tokens against 131K.

How many benchmarks do Granite 4.0 Micro and Muse Spark 1.3 share?

2 benchmarks have published results for both models. Granite 4.0 Micro has 8 scored results on Noometry and Muse Spark 1.3 has 37.

Related comparisons

Go deeper