Model comparison

Gemini 2.0 Pro vs Gemini 2.5 Pro

Gemini 2.5 Pro is the stronger model overall, scoring 45.0 to 39.1 on the Noometry Index.

Last verified . 14 shared benchmarks.

Gemini 2.0 Pro Google

39.1

Rank #173 Confirmed

Gemini 2.5 Pro Google

45.0

Rank #75 Confirmed

Summary

  • They share 14 benchmarks with published results for both. Gemini 2.0 Pro scores higher in 2 categories and Gemini 2.5 Pro in 5 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in long context, where Gemini 2.5 Pro leads 59.8 to 29.2.
  • The biggest single-benchmark swing is Fiction.LiveBench: 41.7% for Gemini 2.0 Pro and 91.7% for Gemini 2.5 Pro.

Side by side

Gemini 2.0 Pro and Gemini 2.5 Pro specifications
Gemini 2.0 ProGemini 2.5 Pro
ProviderGoogleGoogle
Noometry Index39.145.0
Released2025-02-052025-03-25
WeightsProprietaryProprietary
Context window—1.05M
Max output—66K
Input $ / M tokens—$1.25
Output $ / M tokens—$10
Results tracked1478

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Gemini 2.5 Pro leads

Gemini 2.0 Pro: 37.8 (#187), Gemini 2.5 Pro: 42.4 (#101)

Coding benchmarks
BenchmarkGemini 2.0 ProGemini 2.5 Pro
Aider Polyglot35.6%83.1%
LiveBench Coding63.5%85.9%
SWE-bench Verified—57.6%
SWE-bench Verified (bash only)—53.6%
LMArena WebDev—1227
SciCode—42.8%
GSO—3.9%
WeirdML—54%
LMArena Coding—1452
CadEval—64%
ALE-Bench—785.52
AlgoTune—1.51

Agentic & Tool Use Not comparable

Gemini 2.0 Pro: —, Gemini 2.5 Pro: 29.2 (#88)

Agentic & Tool Use benchmarks
BenchmarkGemini 2.0 ProGemini 2.5 Pro
Terminal-Bench—32.6%
GDPval—23.3%
Remote Labor Index—0.8%
TheAgentCompany—30.3%
τ²-bench Banking—13.7%
DeepResearch Bench—42.8%
BALROG—43.3%
LMArena Search—1142
METR Time Horizons—55.4%
Vending-Bench 2—573.64

Reasoning Gemini 2.5 Pro leads

Gemini 2.0 Pro: 22.3 (#198), Gemini 2.5 Pro: 28.8 (#99)

Reasoning benchmarks
BenchmarkGemini 2.0 ProGemini 2.5 Pro
EnigmaEval0.7%5.6%
LiveBench Reasoning60.1%89.8%
LiveBench Data Analysis68%79.9%
Epoch Capabilities Index135.06145.32
LiveBench65.1%82.3%
ARC-AGI-2—4.9%
SimpleBench—62.4%
Kagi LLM Benchmark—70.3%
ARC-AGI-1—41%
CritPt—2%
Chess Puzzles—20%
LMArena Hard Prompts—1455
DTBench—82.4%
LMCA—34.8%
ForecastBench—61.3

Math Gemini 2.0 Pro leads

Gemini 2.0 Pro: 39.7 (#100), Gemini 2.5 Pro: 32.5 (#213)

Math benchmarks
BenchmarkGemini 2.0 ProGemini 2.5 Pro
LiveBench Math71%90.2%
MATH Level 583.5%95.9%
FrontierMath (Tiers 1-3)—24.6%
FrontierMath Tier 4—0%
OTIS Mock AIME 2024-2025—84.7%
Omni-MATH—41.6%
LMArena Math—1450
FrontierMath (Feb 2025 set)—14.1%
FrontierMath Tier 4 (v1)—4.2%

Knowledge Gemini 2.5 Pro leads

Gemini 2.0 Pro: 36.5 (#167), Gemini 2.5 Pro: 56.0 (#46)

Knowledge benchmarks
BenchmarkGemini 2.0 ProGemini 2.5 Pro
GPQA Diamond65.7%85.3%
Confabulations18.4%10.6%
Humanity's Last Exam—21.6%
MMLU-Pro—86.3%
Vectara Hallucination Rate—7%
GPQA (HELM)—74.9%
LMArena Expert—1452

Multimodal Not comparable

Gemini 2.0 Pro: —, Gemini 2.5 Pro: 45.2 (#18)

Multimodal benchmarks
BenchmarkGemini 2.0 ProGemini 2.5 Pro
LMArena Vision—1263
GeoBench—86%
VPCT—48%
LMArena Document—1421
SpatialViz-Bench—44.7%

Multilingual Not comparable

Gemini 2.0 Pro: —, Gemini 2.5 Pro: 55.3 (#31)

Multilingual benchmarks
BenchmarkGemini 2.0 ProGemini 2.5 Pro
LMArena Non-English—1451
LMArena Chinese—1507
LMArena French—1472
LMArena German—1487
LMArena Japanese—1461
LMArena Korean—1434
LMArena Russian—1461
LMArena Spanish—1473

Instruction Following Too close to call

Gemini 2.0 Pro: 75.5 (#59), Gemini 2.5 Pro: 75.0 (#75)

Instruction Following benchmarks
BenchmarkGemini 2.0 ProGemini 2.5 Pro
LiveBench Instruction Following83.4%80.6%
IFEval—84%
LMArena Instruction Following—1437

Long Context Gemini 2.5 Pro leads

Gemini 2.0 Pro: 29.2 (#292), Gemini 2.5 Pro: 59.8 (#5)

Long Context benchmarks
BenchmarkGemini 2.0 ProGemini 2.5 Pro
Fiction.LiveBench41.7%91.7%
LMArena Longer Query—1449

Writing & Preference Gemini 2.5 Pro leads

Gemini 2.0 Pro: 52.7 (#165), Gemini 2.5 Pro: 63.7 (#62)

Writing & Preference benchmarks
BenchmarkGemini 2.0 ProGemini 2.5 Pro
LiveBench Language44.9%67.8%
LMArena Text—1458
LMArena Creative Writing—1454
Short-Story Creative Writing—83.8%
EQ-Bench Creative Writing—1421
WildBench—85.7%
LMArena Multi-Turn—1453

Frequently asked questions

Is Gemini 2.0 Pro better than Gemini 2.5 Pro?

Gemini 2.5 Pro is the stronger model overall, scoring 45.0 to 39.1 on the Noometry Index.

Is Gemini 2.0 Pro or Gemini 2.5 Pro better for coding?

Gemini 2.5 Pro scores higher on coding benchmarks: 42.4 versus 37.8 in the Noometry coding category.

How many benchmarks do Gemini 2.0 Pro and Gemini 2.5 Pro share?

14 benchmarks have published results for both models. Gemini 2.0 Pro has 14 scored results on Noometry and Gemini 2.5 Pro has 78.

Related comparisons

Go deeper