Model comparison

Claude Fable 5 vs GPT-5.5

Claude Fable 5 is the stronger model overall, scoring 66.8 to 63.4 on the Noometry Index. GPT-5.5 costs 1.8× less per token, which makes it the better buy when Claude Fable 5's lead doesn't matter for your workload.

Last verified . 60 shared benchmarks.

Claude Fable 5 Anthropic

66.8

Rank #5 Confirmed

GPT-5.5 OpenAI

63.4

Rank #9 Confirmed

Summary

  • They share 60 benchmarks with published results for both. Claude Fable 5 scores higher in 7 categories and GPT-5.5 in 3 categories; 9 gaps are clear of the uncertainty.
  • The widest gap is in coding, where Claude Fable 5 leads 70.6 to 58.2.
  • The biggest single-benchmark swing is MirrorCode: 63.9% for Claude Fable 5 and 10% for GPT-5.5.
  • GPT-5.5 is cheaper at $5 / $30 per million input/output tokens, against $10 / $50 for Claude Fable 5.
  • GPT-5.5 accepts more context: 1.05M tokens versus 1M.

Side by side

Claude Fable 5 and GPT-5.5 specifications
Claude Fable 5GPT-5.5
ProviderAnthropicOpenAI
Noometry Index66.863.4
Released2026-06-072026-04-23
WeightsProprietaryProprietary
Context window1M1.05M
Max output128K128K
Input $ / M tokens$10$5
Output $ / M tokens$50$30
Results tracked6271

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Claude Fable 5 leads

Claude Fable 5: 70.6 (#4), GPT-5.5: 58.2 (#17)

Coding benchmarks
BenchmarkClaude Fable 5GPT-5.5
DeepSWE69.9%67%
FrontierCode53.5%43%
LMArena WebDev16251513
SciCode61%56.1%
GSO78.4%40.2%
WeirdML91.9%84.9%
LMArena Coding15191494
MirrorCode63.9%10%
ALE-Bench2,0411,943
SWE-bench Verified—80.6%
FrontierSWE47%—

Agentic & Tool Use Claude Fable 5 leads

Claude Fable 5: 54.0 (#2), GPT-5.5: 50.7 (#6)

Agentic & Tool Use benchmarks
BenchmarkClaude Fable 5GPT-5.5
APEX-Agents63.6%55.1%
Remote Labor Index16.1%6.3%
τ²-bench Banking39.7%44.6%
PostTrainBench41.8%27.2%
GBAEval74.5%53.2%
GDP.pdf30%26%
LMArena Search12301242
Vending-Bench 25,6807,524
Terminal-Bench—84.7%
OSWorld 2.0—13%
DeepResearch Bench—54%
ExploitBench—47.4%

Reasoning Claude Fable 5 leads

Claude Fable 5: 76.8 (#6), GPT-5.5: 72.8 (#11)

Reasoning benchmarks
BenchmarkClaude Fable 5GPT-5.5
ARC-AGI-289.2%85%
SimpleBench81.9%69%
Kagi LLM Benchmark91.4%88.8%
NYT Connections (extended)92.7%96.2%
ARC-AGI-198.5%95%
CritPt28.6%27.1%
Chess Puzzles41%54%
EBR-Bench39.5%34.3%
LMArena Hard Prompts15081489
Mystery Game Puzzles52%56%
DTBench98.4%96%
LMCA61.1%54.3%
Surface Evolver Bench95%88.1%
Bench to the Future 30.130.14
Epoch Capabilities Index162.06159.1
EnigmaEval39.3%—
ForecastBench—60.6

Math Claude Fable 5 leads

Claude Fable 5: 88.5 (#5), GPT-5.5: 81.7 (#11)

Knowledge GPT-5.5 leads

Claude Fable 5: 62.2 (#25), GPT-5.5: 64.4 (#17)

Knowledge benchmarks
BenchmarkClaude Fable 5GPT-5.5
GPQA Diamond85.9%94%
SimpleQA Verified70.7%63%
LMArena Expert15341508
Vectara Hallucination Rate—9.3%

Multimodal GPT-5.5 leads

Claude Fable 5: 45.3 (#17), GPT-5.5: 46.9 (#12)

Multimodal benchmarks
BenchmarkClaude Fable 5GPT-5.5
LMArena Vision13241297
Blueprint-Bench 238.6%36.2%
Furniture Assembly35.8%44.2%
LMArena Document14961486

Multilingual Too close to call

Claude Fable 5: 57.3 (#9), GPT-5.5: 56.4 (#20)

Multilingual benchmarks
BenchmarkClaude Fable 5GPT-5.5
LMArena Non-English14811467
LMArena Chinese15431533
LMArena French15051486
LMArena German14861480
LMArena Japanese15061498
LMArena Korean14881460
LMArena Russian15041473
LMArena Spanish14981468

Instruction Following Claude Fable 5 leads

Claude Fable 5: 78.6 (#8), GPT-5.5: 77.5 (#18)

Instruction Following benchmarks
BenchmarkClaude Fable 5GPT-5.5
LMArena Instruction Following15021479

Long Context GPT-5.5 leads

Claude Fable 5: 46.3 (#23), GPT-5.5: 48.3 (#12)

Long Context benchmarks
BenchmarkClaude Fable 5GPT-5.5
LMArena Longer Query15091484
CL-bench Life—22.2%

Writing & Preference Claude Fable 5 leads

Claude Fable 5: 75.9 (#5), GPT-5.5: 72.7 (#13)

Writing & Preference benchmarks
BenchmarkClaude Fable 5GPT-5.5
LMArena Text14911472
LMArena Creative Writing14941455
EQ-Bench Creative Writing19431844
EQ-Bench 413401315
LMArena Multi-Turn15041476

Frequently asked questions

Is Claude Fable 5 better than GPT-5.5?

Claude Fable 5 is the stronger model overall, scoring 66.8 to 63.4 on the Noometry Index. GPT-5.5 costs 1.8× less per token, which makes it the better buy when Claude Fable 5's lead doesn't matter for your workload.

Which is cheaper, Claude Fable 5 or GPT-5.5?

GPT-5.5 is cheaper. It lists at $5 per million input tokens and $30 per million output tokens; Claude Fable 5 lists at $10 and $50.

Is Claude Fable 5 or GPT-5.5 better for coding?

Claude Fable 5 scores higher on coding benchmarks: 70.6 versus 58.2 in the Noometry coding category.

Which has the bigger context window?

GPT-5.5 does, with 1.05M tokens against 1M.

How many benchmarks do Claude Fable 5 and GPT-5.5 share?

60 benchmarks have published results for both models. Claude Fable 5 has 62 scored results on Noometry and GPT-5.5 has 71.

Related comparisons

Go deeper