Model comparison

Gemini 2.0 Pro vs Muse Spark 1.3

Muse Spark 1.3 is the stronger model overall, scoring 54.8 to 39.1 on the Noometry Index.

Last verified . 1 shared benchmarks.

Gemini 2.0 Pro Google

39.1

Rank #173 Confirmed

Muse Spark 1.3 Meta

54.8

Rank #27 Confirmed

Summary

  • They share 1 benchmark with published results for both. Gemini 2.0 Pro scores higher in 0 categories and Muse Spark 1.3 in 7 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in math, where Muse Spark 1.3 leads 73.1 to 39.7.

Side by side

Gemini 2.0 Pro and Muse Spark 1.3 specifications
Gemini 2.0 ProMuse Spark 1.3
ProviderGoogleMeta
Noometry Index39.154.8
Released2025-02-052026-09-02
WeightsProprietaryProprietary
Context window—1.05M
Max output—131K
Input $ / M tokens—$1.25
Output $ / M tokens—$4.25
Results tracked1437

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Muse Spark 1.3 leads

Gemini 2.0 Pro: 37.8 (#187), Muse Spark 1.3: 56.6 (#21)

Coding benchmarks
BenchmarkGemini 2.0 ProMuse Spark 1.3
Aider Polyglot35.6%—
CursorBench—41.6%
LMArena WebDev—1657
SciCode—59.7%
LiveBench Coding63.5%—
LMArena Coding—1514

Agentic & Tool Use Not comparable

Gemini 2.0 Pro: —, Muse Spark 1.3: 38.6 (#30)

Agentic & Tool Use benchmarks
BenchmarkGemini 2.0 ProMuse Spark 1.3
APEX-Agents—57.8%
GDP.pdf—27.6%

Reasoning Muse Spark 1.3 leads

Gemini 2.0 Pro: 22.3 (#198), Muse Spark 1.3: 54.0 (#27)

Reasoning benchmarks
BenchmarkGemini 2.0 ProMuse Spark 1.3
Epoch Capabilities Index135.06156.75
NYT Connections (extended)—85.1%
CritPt—26%
Chess Puzzles—38%
EnigmaEval0.7%—
LiveBench Reasoning60.1%—
LMArena Hard Prompts—1503
Mystery Game Puzzles—25%
DTBench—96.5%
LiveBench Data Analysis68%—
LMCA—53.9%
Bench to the Future 3—0.14
LiveBench65.1%—

Math Muse Spark 1.3 leads

Gemini 2.0 Pro: 39.7 (#100), Muse Spark 1.3: 73.1 (#21)

Math benchmarks
BenchmarkGemini 2.0 ProMuse Spark 1.3
FrontierMath (Tiers 1-3)—74.4%
FrontierMath Tier 4—46.3%
OTIS Mock AIME 2024-2025—99.2%
ProofBench—58%
LiveBench Math71%—
LMArena Math—1494
MATH Level 583.5%—

Knowledge Muse Spark 1.3 leads

Gemini 2.0 Pro: 36.5 (#167), Muse Spark 1.3: 42.6 (#95)

Knowledge benchmarks
BenchmarkGemini 2.0 ProMuse Spark 1.3
GPQA Diamond65.7%—
Confabulations18.4%—
LMArena Expert—1516

Multimodal Not comparable

Gemini 2.0 Pro: —, Muse Spark 1.3: 43.7 (#22)

Multimodal benchmarks
BenchmarkGemini 2.0 ProMuse Spark 1.3
LMArena Vision—1309
LMArena Document—1471

Multilingual Not comparable

Gemini 2.0 Pro: —, Muse Spark 1.3: 57.4 (#8)

Multilingual benchmarks
BenchmarkGemini 2.0 ProMuse Spark 1.3
LMArena Non-English—1481
LMArena Chinese—1529
LMArena French—1524
LMArena German—1515
LMArena Japanese—1474
LMArena Korean—1501
LMArena Russian—1490
LMArena Spanish—1490

Instruction Following Muse Spark 1.3 leads

Gemini 2.0 Pro: 75.5 (#59), Muse Spark 1.3: 77.5 (#22)

Instruction Following benchmarks
BenchmarkGemini 2.0 ProMuse Spark 1.3
LiveBench Instruction Following83.4%—
LMArena Instruction Following—1477

Long Context Muse Spark 1.3 leads

Gemini 2.0 Pro: 29.2 (#292), Muse Spark 1.3: 45.6 (#32)

Long Context benchmarks
BenchmarkGemini 2.0 ProMuse Spark 1.3
Fiction.LiveBench41.7%—
LMArena Longer Query—1488

Writing & Preference Muse Spark 1.3 leads

Gemini 2.0 Pro: 52.7 (#165), Muse Spark 1.3: 73.6 (#9)

Writing & Preference benchmarks
BenchmarkGemini 2.0 ProMuse Spark 1.3
LMArena Text—1490
LMArena Creative Writing—1455
EQ-Bench Creative Writing—1906
LMArena Multi-Turn—1482
LiveBench Language44.9%—

Frequently asked questions

Is Gemini 2.0 Pro better than Muse Spark 1.3?

Muse Spark 1.3 is the stronger model overall, scoring 54.8 to 39.1 on the Noometry Index.

Is Gemini 2.0 Pro or Muse Spark 1.3 better for coding?

Muse Spark 1.3 scores higher on coding benchmarks: 56.6 versus 37.8 in the Noometry coding category.

How many benchmarks do Gemini 2.0 Pro and Muse Spark 1.3 share?

1 benchmark has published results for both models. Gemini 2.0 Pro has 14 scored results on Noometry and Muse Spark 1.3 has 37.

Related comparisons

Go deeper