OpenAI, proprietary

GPT-6.1 Sol

GPT-6.1 Sol by OpenAI ranks 6th of 354 ranked models on the Noometry Index as of October 2026, with a score of 65.6. Its strongest category is math, where it ranks 1st. API pricing starts at $2 per million input tokens and $10 per million output tokens, with a 1.05M-token context window.

Last verified

Specifications

Noometry rank
#6 of 354
Index score
65.6
Evidence
Confirmed 34 results
Provider
OpenAI
Released
September 29, 2026
Weights
Proprietary
Reasoning
Yes
Context window
1.05M
Max output
128K
Input price
$2 / M
Output price
$10 / M
Blended price
$4 / M
Output speed
Not measured
Value
#160 of 219
Knowledge cutoff
April 2026
Input
text, image, pdf

Category scores

Each category score combines every public result we have in that category.

GPT-6.1 Sol category scores
  1. Coding 63.2
  2. Agentic & Tool Use 39.6
  3. Reasoning 81.9
  4. Math 93.7
  5. Knowledge 71.8
  6. Multimodal 52.7
  7. Multilingual 54.3
  8. Instruction Following 77.0
  9. Long Context 44.9
  10. Writing & Preference 63.6
GPT-6.1 Sol category ranks
CategoryScoreRankResults
Coding63.2#85
Agentic & Tool Use39.6#262
Reasoning81.9#28
Math93.7#15
Knowledge71.8#43
Multimodal52.7#52
Multilingual54.3#461
Instruction Following77.0#291
Long Context44.9#541
Writing & Preference63.6#633

Strengths and weaknesses

Categories where GPT-6.1 Sol places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

GPT-6.1 Sol: strongest categories
CategoryScorevs medianRank
Math93.7+57.2#1 of 327, top 1%
Reasoning81.9+58.3#2 of 350, top 1%
Knowledge71.8+34.4#4 of 314, top 2%

Weakest categories

GPT-6.1 Sol: weakest categories
CategoryScorevs medianRank
Writing & Preference63.6+9.8#63 of 312, top 21%
Long Context44.9+3.9#54 of 296, top 19%
Agentic & Tool Use39.6+9.3#26 of 154, top 17%

Closest competitors

The models ranked just above and below GPT-6.1 Sol. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to GPT-6.1 Sol
ModelRankScoreBlended $/MSpeed
Claude Fable 5.1#269.0$20—Compare
Claude Opus 5.5#368.6$8—Compare
Claude Opus 5#467.8$10—Compare
Claude Fable 5#566.8$2025Compare
GPT-5.6 Sol#765.0$810Compare
GPT-5.5 Pro#864.3$67.50—Compare
GPT-5.5#963.4$11.2525Compare
Claude Sonnet 5.5#1061.9$4—Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

GPT-6.1 Sol Coding benchmark results
BenchmarkScorePositionSettingSourceDate
DeepSWE75.2%Best of 29highModel card (self-reported)2026-09-29
DeepSWE64.4%lowModel card (self-reported)2026-09-29
DeepSWE71.9%maxModel card (self-reported)2026-09-29
DeepSWE73%mediumModel card (self-reported)2026-09-29
DeepSWE71.9%xhighModel card (self-reported)2026-09-29
FrontierCode50.2%#7 of 37, top 19%mediumEpoch AI
LMArena WebDev1755#4 of 113, top 4%LMArena2026-10-08
SciCode55.8%#24 of 121, top 20%highEpoch AI
SciCode53.2%lowEpoch AI
SciCode54.2%maxEpoch AI
SciCode53.2%mediumEpoch AI
SciCode55.7%xhighEpoch AI
LMArena Coding1487#36 of 294, top 13%LMArena2026-10-08

Agentic & Tool Use

GPT-6.1 Sol Agentic & Tool Use benchmark results
BenchmarkScorePositionSettingSourceDate
APEX-Agents60%#12 of 49, top 25%maxEpoch AI
GDP.pdf32%#2 of 36, top 6%highModel card (self-reported)2026-09-29
GDP.pdf27%lowModel card (self-reported)2026-09-29
GDP.pdf31%maxModel card (self-reported)2026-09-29
GDP.pdf30%mediumModel card (self-reported)2026-09-29
GDP.pdf31.8%xhighModel card (self-reported)2026-09-29

Reasoning

GPT-6.1 Sol Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
ARC-AGI-291.7%highEpoch AI
ARC-AGI-276.7%lowEpoch AI
ARC-AGI-294.2%#2 of 83, top 3%maxEpoch AI
ARC-AGI-286.7%mediumEpoch AI
ARC-AGI-291.7%xhighEpoch AI
NYT Connections (extended)95.5%#5 of 91, top 6%high reasoningLech Mazur benchmarks
ARC-AGI-198.5%#4 of 83, top 5%highEpoch AI
ARC-AGI-193.5%lowEpoch AI
ARC-AGI-196.5%maxEpoch AI
ARC-AGI-195.5%mediumEpoch AI
ARC-AGI-198.5%xhighEpoch AI
CritPt30%highEpoch AI
CritPt24.9%lowEpoch AI
CritPt31.7%#3 of 134, top 3%maxEpoch AI
CritPt27.7%mediumEpoch AI
CritPt31.7%xhighEpoch AI
Chess Puzzles61%#5 of 129, top 4%maxEpoch AI2026-09-29
EBR-Bench54.3%#4 of 24, top 17%maxEpoch AI2026-09-29
LMArena Hard Prompts1466#40 of 297, top 14%LMArena2026-10-08
Mystery Game Puzzles80%#2 of 74, top 3%maxEpoch AI2026-09-29
Epoch Capabilities Index166.09#3 of 213, top 2%Epoch AI2026-09-29

Math

GPT-6.1 Sol Math benchmark results
BenchmarkScorePositionSettingSourceDate
FrontierMath (Tiers 1-3)93.7%Best of 81maxEpoch AI2026-09-29
FrontierMath Tier 4100%Best of 63maxEpoch AI2026-09-29
OTIS Mock AIME 2024-2025100%#8 of 173, top 5%maxEpoch AI2026-09-29
ProofBench99%#6 of 77, top 8%Epoch AI
LMArena Math1464#49 of 285, top 18%LMArena2026-10-08

Knowledge

GPT-6.1 Sol Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
GPQA Diamond95.4%#4 of 186, top 3%maxEpoch AI2026-09-29
SimpleQA Verified73.9%#2 of 77, top 3%maxEpoch AI2026-09-29
LMArena Expert1502#24 of 273, top 9%LMArena2026-10-08

Multimodal

GPT-6.1 Sol Multimodal benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Vision1288#23 of 122, top 19%LMArena2026-10-09
Furniture Assembly80%#2 of 31, top 7%maxEpoch AI2026-09-29

Multilingual

GPT-6.1 Sol Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1438#46 of 297, top 16%LMArena2026-10-08
LMArena Chinese1477#68 of 285, top 24%LMArena2026-10-08
LMArena Russian1455#38 of 283, top 14%LMArena2026-10-08

Instruction Following

GPT-6.1 Sol Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Instruction Following1468#25 of 298, top 9%LMArena2026-10-08

Long Context

GPT-6.1 Sol Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Longer Query1465#36 of 291, top 13%LMArena2026-10-08

Writing & Preference

GPT-6.1 Sol Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1447#50 of 297, top 17%LMArena2026-10-08
LMArena Creative Writing1432#48 of 295, top 17%LMArena2026-10-08
LMArena Multi-Turn1449#56 of 295, top 19%LMArena2026-10-08

API pricing by provider

GPT-6.1 Sol API prices
RouteInput $/MOutput $/MCached input $/MChecked
azure$2$10$0.102026-10-10
bedrock$2$10$0.102026-10-10
openai$2$10$0.102026-10-10
openrouter$2$10$0.102026-10-10

Compare GPT-6.1 Sol

Other OpenAI models

Frequently asked questions

How good is GPT-6.1 Sol?

GPT-6.1 Sol by OpenAI ranks 6th of 354 ranked models on the Noometry Index as of October 2026, with a score of 65.6. Its strongest category is math, where it ranks 1st. API pricing starts at $2 per million input tokens and $10 per million output tokens, with a 1.05M-token context window.

How much does GPT-6.1 Sol cost?

GPT-6.1 Sol costs $2 per million input tokens and $10 per million output tokens on OpenAI's own API, with cached input at $0.10.

What is GPT-6.1 Sol's context window?

GPT-6.1 Sol accepts up to 1.05M tokens of input and can write up to 128K tokens in one response.

Is GPT-6.1 Sol open source?

No. GPT-6.1 Sol is proprietary and available only through OpenAI's API and partner platforms.

What are GPT-6.1 Sol's strengths and weaknesses?

Relative to other ranked models, GPT-6.1 Sol places best in math, reasoning, knowledge and lowest in writing & preference, long context, agentic & tool use.

What is GPT-6.1 Sol best at?

Its best category is math, where it ranks 1st on Noometry.