Coding benchmark
CursorBench leaderboard
As of October 2026, Claude Opus 5.5 has the highest published CursorBench score on Noometry at 57.8%, out of 14 models with results.
Last verified
About CursorBench
Coding-agent tasks drawn from real Cursor sessions, published by Cursor.
- Category
- Coding
- Introduced
- 2026
- Unit
- Percent (random guessing ≈ 0%)
- Official site
- cursor.com
Top 14 models
- Claude Opus 5.5 57.8%
- Claude Sonnet 5.5 55.5%
- Claude Fable 5.1 51.8%
- Claude Opus 5 46.6%
- Grok 4.7 46.3%
- GLM-5.3 42.6%
- GPT-5.6 Sol 41.7%
- Muse Spark 1.3 41.6%
- Grok 4.6 41.4%
- GPT-5.6 Terra 41.3%
- Gemini 3.8 Flash 39.6%
- GLM-5.3-Flash 36.8%
- GPT-5.6 Luna 35.9%
- Claude Sonnet 5 34.1%
Sponsored placements are available on pages like this one. Advertise on Noometry
All results
| # | Model | Provider | Score | Setting | Source | Date |
|---|---|---|---|---|---|---|
| 1 | Claude Opus 5.5 | Anthropic | 57.8% | max | Epoch AI | |
| 2 | Claude Sonnet 5.5 | Anthropic | 55.5% | max | Epoch AI | |
| 3 | Claude Fable 5.1 | Anthropic | 51.8% | max | Epoch AI | |
| 4 | Claude Opus 5 | Anthropic | 46.6% | max | Epoch AI | |
| 5 | Grok 4.7 | xAI | 46.3% | xhigh | Epoch AI | |
| 6 | GLM-5.3 | Z.ai (Zhipu) | 42.6% | max | Epoch AI | |
| 7 | GPT-5.6 Sol | OpenAI | 41.7% | max | Epoch AI | |
| 8 | Muse Spark 1.3 | 41.6% | max | Epoch AI | ||
| 9 | Grok 4.6 | xAI | 41.4% | xhigh | Epoch AI | |
| 10 | GPT-5.6 Terra | OpenAI | 41.3% | max | Epoch AI | |
| 11 | Gemini 3.8 Flash | 39.6% | high | Epoch AI | ||
| 12 | GLM-5.3-Flash | Z.ai (Zhipu) | 36.8% | max | Epoch AI | |
| 13 | GPT-5.6 Luna | OpenAI | 35.9% | max | Epoch AI | |
| 14 | Claude Sonnet 5 | Anthropic | 34.1% | max | Epoch AI |
Compare the leaders
Frequently asked questions
What does CursorBench measure?
Coding-agent tasks drawn from real Cursor sessions, published by Cursor.
Which model has the highest CursorBench score?
As of October 2026, Claude Opus 5.5 has the highest published CursorBench score on Noometry at 57.8%, out of 14 models with results.
What is the best open-weight model on CursorBench?
GLM-5.3 has the highest CursorBench accuracy among open-weight models at 42.6%, ranking 6 of 14 overall.