Anthropic, proprietary
Claude Sonnet 5.5
Claude Sonnet 5.5 by Anthropic ranks 10th of 354 ranked models on the Noometry Index as of October 2026, with a score of 61.9. Its strongest category is coding, where it ranks 6th. API pricing starts at $2 per million input tokens and $10 per million output tokens, with a 1M-token context window.
Last verified
Specifications
- Noometry rank
- #10 of 354
- Index score
- 61.9
- Evidence
- Confirmed 32 results
- Provider
- Anthropic
- Released
- September 28, 2026
- Weights
- Proprietary
- Reasoning
- Yes
- Context window
- 1M
- Max output
- 128K
- Input price
- $2 / M
- Output price
- $10 / M
- Blended price
- $4 / M
- Output speed
- Not measured
- Value
- #162 of 219
- Knowledge cutoff
- June 2026
- Input
- text, image, pdf
Category scores
Each category score combines every public result we have in that category.
- Coding 67.3
- Agentic & Tool Use 45.0
- Reasoning 54.0
- Math 87.9
- Knowledge 66.0
- Multimodal 51.5
- Multilingual 55.3
- Instruction Following 78.3
- Long Context 45.9
- Writing & Preference 66.0
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 67.3 | #6 | 6 |
| Agentic & Tool Use | 45.0 | #16 | 1 |
| Reasoning | 54.0 | #28 | 4 |
| Math | 87.9 | #6 | 5 |
| Knowledge | 66.0 | #12 | 3 |
| Multimodal | 51.5 | #6 | 2 |
| Multilingual | 55.3 | #30 | 1 |
| Instruction Following | 78.3 | #11 | 1 |
| Long Context | 45.9 | #28 | 1 |
| Writing & Preference | 66.0 | #40 | 3 |
Strengths and weaknesses
Categories where Claude Sonnet 5.5 places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Coding | 67.3 | +28.5 | #6 of 340, top 2% |
| Math | 87.9 | +51.4 | #6 of 327, top 2% |
| Instruction Following | 78.3 | +7.0 | #11 of 305, top 4% |
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Writing & Preference | 66.0 | +12.2 | #40 of 312, top 13% |
| Agentic & Tool Use | 45.0 | +14.7 | #16 of 154, top 11% |
| Multilingual | 55.3 | +7.9 | #30 of 297, top 11% |
Closest competitors
The models ranked just above and below Claude Sonnet 5.5. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| GPT-6.1 Sol | #6 | 65.6 | $4 | — | Compare |
| GPT-5.6 Sol | #7 | 65.0 | $8 | 10 | Compare |
| GPT-5.5 Pro | #8 | 64.3 | $67.50 | — | Compare |
| GPT-5.5 | #9 | 63.4 | $11.25 | 25 | Compare |
| Gemini 3.8 Flash | #11 | 61.8 | $1.50 | — | Compare |
| GPT-6 Sol | #12 | 61.8 | $4 | — | Compare |
| Claude Opus 4.8 | #13 | 60.7 | $10 | 34 | Compare |
| Gemini 3.7 Flash | #14 | 59.8 | $1.50 | — | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| FrontierCode | 52.1% | #5 of 37, top 14% | xhigh | Epoch AI | |
| FrontierCode | 46.2% | max | Model card (self-reported) | 2026-09-28 | |
| CursorBench | 47.8% | high | Epoch AI | ||
| CursorBench | 35.8% | low | Epoch AI | ||
| CursorBench | 55.5% | #2 of 14, top 15% | max | Epoch AI | |
| CursorBench | 39.2% | medium | Epoch AI | ||
| CursorBench | 53.1% | xhigh | Epoch AI | ||
| LMArena WebDev | 1774 | #3 of 113, top 3% | xhigh | LMArena | 2026-10-08 |
| FrontierSWE | 61.9% | #3 of 18, top 17% | max | Epoch AI | |
| SciCode | 53.7% | high | Epoch AI | ||
| SciCode | 49.1% | low | Epoch AI | ||
| SciCode | 61% | #5 of 121, top 5% | max | Epoch AI | |
| SciCode | 52.9% | medium | Epoch AI | ||
| SciCode | 57.3% | xhigh | Epoch AI | ||
| LMArena Coding | 1513 | #10 of 294, top 4% | xhigh | LMArena | 2026-10-08 |
| ALE-Bench | 1,819 | #10 of 105, top 10% | high | Epoch AI |
Agentic & Tool Use
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| APEX-Agents | 75.5% | #2 of 49, top 5% | max | Epoch AI | |
| APEX-Agents | 44.6% | medium | Epoch AI |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| NYT Connections (extended) | 80.5% | #34 of 91, top 38% | high reasoning | Lech Mazur benchmarks | |
| CritPt | 24.6% | high | Epoch AI | ||
| CritPt | 11.4% | low | Epoch AI | ||
| CritPt | 31.4% | #5 of 134, top 4% | max | Epoch AI | |
| CritPt | 16.9% | medium | Epoch AI | ||
| CritPt | 31.1% | xhigh | Epoch AI | ||
| LMArena Hard Prompts | 1495 | #13 of 297, top 5% | xhigh | LMArena | 2026-10-08 |
| Mystery Game Puzzles | 65% | #4 of 74, top 6% | max | Epoch AI | 2026-09-29 |
| Epoch Capabilities Index | 165.03 | #4 of 213, top 2% | Epoch AI | 2026-09-28 |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| FrontierMath (Tiers 1-3) | 88.8% | #7 of 81, top 9% | max | Epoch AI | 2026-09-29 |
| FrontierMath Tier 4 | 80.5% | #8 of 63, top 13% | max | Epoch AI | 2026-09-29 |
| OTIS Mock AIME 2024-2025 | 100% | #4 of 173, top 3% | max | Epoch AI | 2026-09-29 |
| ProofBench | 100% | #3 of 77, top 4% | max | Epoch AI | |
| LMArena Math | 1510 | #7 of 285, top 3% | xhigh | LMArena | 2026-10-08 |
| FrontierMath Erdős | 2.9% | #3 of 7, top 43% | max | Epoch AI | 2026-10-08 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| GPQA Diamond | 95.6% | #2 of 186, top 2% | max | Epoch AI | 2026-09-29 |
| SimpleQA Verified | 46.5% | #33 of 77, top 43% | max | Epoch AI | 2026-09-29 |
| LMArena Expert | 1540 | #6 of 273, top 3% | xhigh | LMArena | 2026-10-08 |
Multimodal
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Vision | 1289 | #22 of 122, top 19% | xhigh | LMArena | 2026-10-09 |
| Furniture Assembly | 75% | #4 of 31, top 13% | max | Epoch AI | 2026-09-29 |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1452 | #30 of 297, top 11% | xhigh | LMArena | 2026-10-08 |
| LMArena Chinese | 1522 | #24 of 285, top 9% | xhigh | LMArena | 2026-10-08 |
| LMArena Russian | 1451 | #41 of 283, top 15% | xhigh | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Instruction Following | 1495 | #8 of 298, top 3% | xhigh | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 1498 | #10 of 291, top 4% | xhigh | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1471 | #23 of 297, top 8% | xhigh | LMArena | 2026-10-08 |
| LMArena Creative Writing | 1465 | #16 of 295, top 6% | xhigh | LMArena | 2026-10-08 |
| LMArena Multi-Turn | 1474 | #26 of 295, top 9% | xhigh | LMArena | 2026-10-08 |
API pricing by provider
| Route | Input $/M | Output $/M | Cached input $/M | Checked |
|---|---|---|---|---|
| anthropic | $2 | $10 | $0.10 | 2026-10-10 |
| azure | $2 | $10 | $0.20 | 2026-10-10 |
| bedrock | $2 | $10 | $0.10 | 2026-10-10 |
| openrouter | $2 | $10 | $0.10 | 2026-10-10 |
| vertex | $2 | $10 | $0.20 | 2026-10-10 |
Compare Claude Sonnet 5.5
- Claude Sonnet 5.5 vs Claude Sonnet 5
- Claude Sonnet 5.5 vs GPT-5.5
- Claude Sonnet 5.5 vs Gemini 3.8 Flash
- Claude Sonnet 5.5 vs GPT-5.5 Pro
- Claude Sonnet 5.5 vs GPT-6 Sol
- Claude Sonnet 5.5 vs GPT-5.6 Sol
- Claude Sonnet 5.5 vs Claude Opus 4.8
- Claude Sonnet 5.5 vs GPT-6 Astra
- Claude Sonnet 5.5 vs Kimi K3
- Claude Sonnet 5.5 vs Grok 4.6
- Claude Sonnet 5.5 vs Qwen3.8 Max
- Claude Sonnet 5.5 vs GLM-5.3
- Claude Sonnet 5.5 vs Muse Spark 1.3
- Claude Sonnet 5.5 vs DeepSeek V4 Pro
Other Anthropic models
- Claude Fable 5.169.0
- Claude Opus 5.568.6
- Claude Opus 567.8
- Claude Fable 566.8
- Claude Opus 4.860.7
- Claude Opus 4.758.3
- Claude Opus 4.658.2
- Claude Sonnet 554.6
Frequently asked questions
How good is Claude Sonnet 5.5?
Claude Sonnet 5.5 by Anthropic ranks 10th of 354 ranked models on the Noometry Index as of October 2026, with a score of 61.9. Its strongest category is coding, where it ranks 6th. API pricing starts at $2 per million input tokens and $10 per million output tokens, with a 1M-token context window.
How much does Claude Sonnet 5.5 cost?
Claude Sonnet 5.5 costs $2 per million input tokens and $10 per million output tokens on Anthropic's own API, with cached input at $0.10.
What is Claude Sonnet 5.5's context window?
Claude Sonnet 5.5 accepts up to 1M tokens of input and can write up to 128K tokens in one response.
Is Claude Sonnet 5.5 open source?
No. Claude Sonnet 5.5 is proprietary and available only through Anthropic's API and partner platforms.
What are Claude Sonnet 5.5's strengths and weaknesses?
Relative to other ranked models, Claude Sonnet 5.5 places best in coding, math, instruction following and lowest in writing & preference, agentic & tool use, multilingual.
What is Claude Sonnet 5.5 best at?
Its best category is coding, where it ranks 6th on Noometry.