Math benchmark
FrontierMath (Feb 2025 set) leaderboard
As of October 2026, GPT-5.5 Pro has the highest published FrontierMath (Feb 2025 set) score on Noometry at 52.4%, out of 68 models with results.
Last verified
About FrontierMath (Feb 2025 set)
Earlier FrontierMath problem set, kept for historical comparison.
Top 15 models
- GPT-5.5 Pro 52.4%
- GPT-5.5 51.7%
- GPT-5.4 Pro 50%
- GPT-5.4 47.6%
- Claude Opus 4.8 47.2%
- Claude Opus 4.7 43.8%
- Claude Opus 4.6 40.7%
- GPT-5.2 40.7%
- Muse Spark 39%
- Gemini 3.5 Flash 39%
- Kimi K2.6 39%
- Gemini 3 Pro 37.6%
- Gemini 3.1 Pro Preview 36.9%
- Gemini 3 Flash Preview 35.6%
- GLM-5.1 33.4%
Sponsored placements are available on pages like this one. Advertise on Noometry
All results
Compare the leaders
Other math benchmarks
Frequently asked questions
What does FrontierMath (Feb 2025 set) measure?
Earlier FrontierMath problem set, kept for historical comparison.
Which model has the highest FrontierMath (Feb 2025 set) score?
As of October 2026, GPT-5.5 Pro has the highest published FrontierMath (Feb 2025 set) score on Noometry at 52.4%, out of 68 models with results.
What is the best open-weight model on FrontierMath (Feb 2025 set)?
Kimi K2.6 has the highest FrontierMath (Feb 2025 set) accuracy among open-weight models at 39%, ranking 11 of 68 overall.