Model comparison

Claude Fable 5.1 vs Gemini 1.5 Pro (May 2024)

Claude Fable 5.1 is the stronger model overall, scoring 69.0 to 32.1 on the Noometry Index.

Last verified . 24 shared benchmarks.

Claude Fable 5.1 Anthropic

69.0

Rank #2 Confirmed

Gemini 1.5 Pro (May 2024) Google

32.1

Rank #261 Confirmed

Summary

  • They share 24 benchmarks with published results for both. Claude Fable 5.1 scores higher in 10 categories and Gemini 1.5 Pro (May 2024) in 0 categories; 10 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Claude Fable 5.1 leads 76.7 to 12.3.
  • The biggest single-benchmark swing is ARC-AGI-2: 90% for Claude Fable 5.1 and 0.8% for Gemini 1.5 Pro (May 2024).

Side by side

Claude Fable 5.1 and Gemini 1.5 Pro (May 2024) specifications
Claude Fable 5.1Gemini 1.5 Pro (May 2024)
ProviderAnthropicGoogle
Noometry Index69.032.1
Released2026-09-012024-02-15
WeightsProprietaryProprietary
Context window1M—
Max output128K—
Input $ / M tokens$10—
Output $ / M tokens$50—
Results tracked5245

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Claude Fable 5.1 leads

Claude Fable 5.1: 74.7 (#1), Gemini 1.5 Pro (May 2024): 34.2 (#241)

Coding benchmarks
BenchmarkClaude Fable 5.1Gemini 1.5 Pro (May 2024)
WeirdML92.9%22.2%
LMArena Coding15281294
FrontierCode50.9%—
CursorBench51.8%—
LMArena WebDev1744—
FrontierSWE56.3%—
SciCode63.1%—
GSO88.2%—
BigCodeBench Instruct—43.8%
MirrorCode73.3%—
BigCodeBench Complete—57.5%
CadEval—34%
ALE-Bench2,143—
HumanEval+—79.3%
MBPP+—74.6%

Agentic & Tool Use Claude Fable 5.1 leads

Claude Fable 5.1: 50.7 (#5), Gemini 1.5 Pro (May 2024): 17.9 (#145)

Agentic & Tool Use benchmarks
BenchmarkClaude Fable 5.1Gemini 1.5 Pro (May 2024)
APEX-Agents68.6%—
Remote Labor Index17.9%—
TheAgentCompany—3.4%
Cybench—7.5%
BALROG—21%
GDP.pdf29.6%—
Vending-Bench 25,422—

Reasoning Claude Fable 5.1 leads

Claude Fable 5.1: 76.7 (#7), Gemini 1.5 Pro (May 2024): 12.3 (#338)

Reasoning benchmarks
BenchmarkClaude Fable 5.1Gemini 1.5 Pro (May 2024)
ARC-AGI-290%0.8%
LMArena Hard Prompts15261296
DTBench97.6%59%
Epoch Capabilities Index164.7131.73
SimpleBench—27.1%
NYT Connections (extended)90%—
ARC-AGI-197.5%—
CritPt31.1%—
Chess Puzzles47%—
EBR-Bench57.1%—
Mystery Game Puzzles58%—
LMCA65.5%—
BIG-Bench Hard—89.2%
ForecastBench—58.4

Math Claude Fable 5.1 leads

Claude Fable 5.1: 89.6 (#4), Gemini 1.5 Pro (May 2024): 25.8 (#266)

Math benchmarks
BenchmarkClaude Fable 5.1Gemini 1.5 Pro (May 2024)
OTIS Mock AIME 2024-2025100%23.1%
LMArena Math15251315
FrontierMath (Tiers 1-3)90.2%—
FrontierMath Tier 487.8%—
ProofBench100%—
Omni-MATH—36.4%
MATH Level 5—70.4%
FrontierMath Erdős0%—

Knowledge Claude Fable 5.1 leads

Claude Fable 5.1: 69.6 (#6), Gemini 1.5 Pro (May 2024): 29.4 (#239)

Knowledge benchmarks
BenchmarkClaude Fable 5.1Gemini 1.5 Pro (May 2024)
Humanity's Last Exam46.5%4.6%
LMArena Expert15351279
GPQA Diamond—57.2%
SimpleQA Verified70.8%—
MMLU-Pro—73.7%
Confabulations—13.5%
GPQA (HELM)—53.4%
MMLU—86.9%

Multimodal Claude Fable 5.1 leads

Claude Fable 5.1: 53.9 (#4), Gemini 1.5 Pro (May 2024): 36.8 (#77)

Multimodal benchmarks
BenchmarkClaude Fable 5.1Gemini 1.5 Pro (May 2024)
LMArena Vision13181161
Video-MME—75%
Blueprint-Bench 241.9%—
Furniture Assembly70%—
LMArena Document1513—

Multilingual Claude Fable 5.1 leads

Claude Fable 5.1: 59.1 (#3), Gemini 1.5 Pro (May 2024): 45.3 (#174)

Multilingual benchmarks
BenchmarkClaude Fable 5.1Gemini 1.5 Pro (May 2024)
LMArena Non-English15071312
LMArena Chinese15861331
LMArena French15251302
LMArena German15001286
LMArena Japanese15431292
LMArena Korean15341298
LMArena Russian15211320
LMArena Spanish15161311

Instruction Following Claude Fable 5.1 leads

Claude Fable 5.1: 79.2 (#6), Gemini 1.5 Pro (May 2024): 68.6 (#185)

Instruction Following benchmarks
BenchmarkClaude Fable 5.1Gemini 1.5 Pro (May 2024)
LMArena Instruction Following15171297
IFEval—83.7%

Long Context Claude Fable 5.1 leads

Claude Fable 5.1: 46.7 (#20), Gemini 1.5 Pro (May 2024): 39.8 (#169)

Long Context benchmarks
BenchmarkClaude Fable 5.1Gemini 1.5 Pro (May 2024)
LMArena Longer Query15221308

Writing & Preference Claude Fable 5.1 leads

Claude Fable 5.1: 79.2 (#2), Gemini 1.5 Pro (May 2024): 52.4 (#172)

Writing & Preference benchmarks
BenchmarkClaude Fable 5.1Gemini 1.5 Pro (May 2024)
LMArena Text15101319
LMArena Creative Writing15071333
LMArena Multi-Turn14921296
EQ-Bench Creative Writing2162—
WildBench—81.3%

Frequently asked questions

Is Claude Fable 5.1 better than Gemini 1.5 Pro (May 2024)?

Claude Fable 5.1 is the stronger model overall, scoring 69.0 to 32.1 on the Noometry Index.

Is Claude Fable 5.1 or Gemini 1.5 Pro (May 2024) better for coding?

Claude Fable 5.1 scores higher on coding benchmarks: 74.7 versus 34.2 in the Noometry coding category.

How many benchmarks do Claude Fable 5.1 and Gemini 1.5 Pro (May 2024) share?

24 benchmarks have published results for both models. Claude Fable 5.1 has 52 scored results on Noometry and Gemini 1.5 Pro (May 2024) has 45.

Related comparisons

Go deeper