Coding benchmark

GSO leaderboard

As of October 2026, Claude Fable 5.1 has the highest published GSO score on Noometry at 88.2%, out of 31 models with results.

Last verified

About GSO

Software optimization tasks: the model must speed up real code bases while keeping them correct.

Category
Coding
Introduced
2025
Format
Performance patch
Unit
Percent (random guessing ≈ 0%)
Official site
gso-bench.github.io

Top 15 models

Top models on GSO
  1. Claude Fable 5.1 88.2%
  2. GPT-6 Astra 79.4%
  3. Claude Fable 5 78.4%
  4. GPT-5.6 Sol 76.5%
  5. Claude Opus 4.8 47.1%
  6. Claude Opus 4.7 44.1%
  7. Claude Opus 4.6 41.2%
  8. GPT-5.5 40.2%
  9. Claude Sonnet 5 37.3%
  10. GPT-5.4 31.4%
  11. GPT-5.2 27.4%
  12. Claude Opus 4.5 26.5%
  13. Gemini 3.1 Pro Preview 22.6%
  14. Gemini 3 Pro 18.6%
  15. Claude Sonnet 4.5 14.7%

Sponsored placements are available on pages like this one. Advertise on Noometry

All results

Compare the leaders

Other coding benchmarks

Frequently asked questions

What does GSO measure?

Software optimization tasks: the model must speed up real code bases while keeping them correct.

Which model has the highest GSO score?

As of October 2026, Claude Fable 5.1 has the highest published GSO score on Noometry at 88.2%, out of 31 models with results.

What is the best open-weight model on GSO?

Kimi K2 (Jul 2025) has the highest GSO accuracy among open-weight models at 4.9%, ranking 22 of 31 overall.