OpenAI, proprietary

GPT-6 Sol

GPT-6 Sol by OpenAI ranks 12th of 354 ranked models on the Noometry Index as of October 2026, with a score of 61.8. Its strongest category is math, where it ranks 7th. API pricing starts at $2 per million input tokens and $10 per million output tokens, with a 1.05M-token context window.

Last verified

Specifications

Noometry rank
#12 of 354
Index score
61.8
Evidence
Confirmed 45 results
Provider
OpenAI
Released
September 22, 2026
Weights
Proprietary
Reasoning
Yes
Context window
1.05M
Max output
128K
Input price
$2 / M
Output price
$10 / M
Blended price
$4 / M
Output speed
Not measured
Value
#163 of 219
Knowledge cutoff
April 2026
Input
text, image, pdf

Category scores

Each category score combines every public result we have in that category.

GPT-6 Sol category scores
  1. Coding 60.1
  2. Agentic & Tool Use 37.2
  3. Reasoning 74.0
  4. Math 87.2
  5. Knowledge 64.8
  6. Multimodal 47.6
  7. Multilingual 50.5
  8. Instruction Following 74.5
  9. Long Context 43.1
  10. Writing & Preference 71.9
GPT-6 Sol category ranks
CategoryScoreRankResults
Coding60.1#115
Agentic & Tool Use37.2#362
Reasoning74.0#99
Math87.2#75
Knowledge64.8#154
Multimodal47.6#103
Multilingual50.5#1181
Instruction Following74.5#941
Long Context43.1#1081
Writing & Preference71.9#184

Strengths and weaknesses

Categories where GPT-6 Sol places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

GPT-6 Sol: strongest categories
CategoryScorevs medianRank
Math87.2+50.7#7 of 327, top 3%
Reasoning74.0+50.4#9 of 350, top 3%
Coding60.1+21.4#11 of 340, top 4%

Weakest categories

GPT-6 Sol: weakest categories
CategoryScorevs medianRank
Multilingual50.5+3.1#118 of 297, top 40%
Long Context43.1+2.1#108 of 296, top 37%
Instruction Following74.5+3.2#94 of 305, top 31%

Closest competitors

The models ranked just above and below GPT-6 Sol. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to GPT-6 Sol
ModelRankScoreBlended $/MSpeed
GPT-5.5 Pro#864.3$67.50—Compare
GPT-5.5#963.4$11.2525Compare
Claude Sonnet 5.5#1061.9$4—Compare
Gemini 3.8 Flash#1161.8$1.50—Compare
Claude Opus 4.8#1360.7$1034Compare
Gemini 3.7 Flash#1459.8$1.50—Compare
Kimi K3#1559.5$6—Compare
GPT-5.4#1659.4$5.6312Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

GPT-6 Sol Coding benchmark results
BenchmarkScorePositionSettingSourceDate
DeepSWE65.3%highModel card (self-reported)2026-09-22
DeepSWE37.2%lowModel card (self-reported)2026-09-22
DeepSWE68.8%#9 of 29, top 32%maxModel card (self-reported)2026-09-22
DeepSWE56.6%mediumModel card (self-reported)2026-09-22
DeepSWE66.6%xhighModel card (self-reported)2026-09-22
FrontierCode49.3%#8 of 37, top 22%maxEpoch AI
FrontierCode47.7%highModel card (self-reported)2026-09-22
FrontierCode37.3%lowModel card (self-reported)2026-09-22
FrontierCode49.3%maxModel card (self-reported)2026-09-22
FrontierCode45.9%mediumModel card (self-reported)2026-09-22
FrontierCode48.4%xhighModel card (self-reported)2026-09-22
LMArena WebDev1688#7 of 113, top 7%LMArena2026-10-08
SciCode54.9%highEpoch AI
SciCode50.2%lowEpoch AI
SciCode57.6%#15 of 121, top 13%maxEpoch AI
SciCode53.8%mediumEpoch AI
SciCode47.3%noneEpoch AI
SciCode55.1%xhighEpoch AI
LMArena Coding1447#89 of 294, top 31%LMArena2026-10-08
ALE-Bench2,462#2 of 105, top 2%maxEpoch AI

Agentic & Tool Use

GPT-6 Sol Agentic & Tool Use benchmark results
BenchmarkScorePositionSettingSourceDate
APEX-Agents54.3%#21 of 49, top 43%maxEpoch AI
GDP.pdf26.4%#8 of 36, top 23%maxEpoch AI
GDP.pdf28%highModel card (self-reported)2026-09-29
GDP.pdf21.8%lowModel card (self-reported)2026-09-29
GDP.pdf24.8%maxModel card (self-reported)2026-09-29
GDP.pdf25.4%mediumModel card (self-reported)2026-09-29
GDP.pdf23.8%xhighModel card (self-reported)2026-09-29
Vending-Bench 214,428#2 of 60, top 4%Epoch AI

Reasoning

GPT-6 Sol Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
ARC-AGI-268.9%highEpoch AI
ARC-AGI-231.5%lowEpoch AI
ARC-AGI-289.6%#7 of 83, top 9%maxEpoch AI
ARC-AGI-257.8%mediumEpoch AI
ARC-AGI-21.7%noneEpoch AI
ARC-AGI-278.1%xhighEpoch AI
NYT Connections (extended)90.1%#17 of 91, top 19%high reasoningLech Mazur benchmarks
ARC-AGI-191%highEpoch AI
ARC-AGI-172.2%lowEpoch AI
ARC-AGI-195.5%#14 of 83, top 17%maxEpoch AI
ARC-AGI-183.7%mediumEpoch AI
ARC-AGI-129.3%noneEpoch AI
ARC-AGI-192.7%xhighEpoch AI
CritPt25.4%highEpoch AI
CritPt16.3%lowEpoch AI
CritPt30.9%#7 of 134, top 6%maxEpoch AI
CritPt24.6%mediumEpoch AI
CritPt4%noneEpoch AI
CritPt28%xhighEpoch AI
EBR-Bench53.3%#5 of 24, top 21%maxEpoch AI2026-09-22
LMArena Hard Prompts1418#107 of 297, top 37%LMArena2026-10-08
Mystery Game Puzzles56%#9 of 74, top 13%maxEpoch AI2026-09-22
DTBench97.3%#6 of 151, top 4%maxEpoch AI
LMCA59.1%#7 of 125, top 6%maxEpoch AI
Epoch Capabilities Index162.72#7 of 213, top 4%Epoch AI2026-09-22

Math

GPT-6 Sol Math benchmark results
BenchmarkScorePositionSettingSourceDate
FrontierMath (Tiers 1-3)89.8%#5 of 81, top 7%maxEpoch AI2026-09-22
FrontierMath Tier 490%#5 of 63, top 8%maxEpoch AI2026-09-22
OTIS Mock AIME 2024-2025100%#10 of 173, top 6%maxEpoch AI2026-09-22
ProofBench83%#11 of 77, top 15%Epoch AI
LMArena Math1402#123 of 285, top 44%LMArena2026-10-08

Knowledge

GPT-6 Sol Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
GPQA Diamond94.3%#8 of 186, top 5%maxEpoch AI2026-09-22
SimpleQA Verified60.7%#14 of 77, top 19%maxEpoch AI2026-09-22
Vectara Hallucination Rate (lower is better)6.5%#25 of 96, top 27%Vectara Hallucination Leaderboard
LMArena Expert1439#86 of 273, top 32%LMArena2026-10-08

Multimodal

GPT-6 Sol Multimodal benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Vision1245#60 of 122, top 50%LMArena2026-10-09
Blueprint-Bench 236.9%#7 of 31, top 23%Epoch AI
Furniture Assembly58.3%#7 of 31, top 23%maxEpoch AI2026-09-23

Multilingual

GPT-6 Sol Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1385#118 of 297, top 40%LMArena2026-10-08
LMArena Chinese1405#135 of 285, top 48%LMArena2026-10-08
LMArena French1410#111 of 223, top 50%LMArena2026-10-08
LMArena German1390#103 of 231, top 45%LMArena2026-10-08
LMArena Japanese1385#72 of 211, top 35%LMArena2026-10-08
LMArena Korean1341#109 of 213, top 52%LMArena2026-10-08
LMArena Russian1401#103 of 283, top 37%LMArena2026-10-08
LMArena Spanish1384#124 of 226, top 55%LMArena2026-10-08

Instruction Following

GPT-6 Sol Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Instruction Following1412#84 of 298, top 29%LMArena2026-10-08

Long Context

GPT-6 Sol Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Longer Query1411#105 of 291, top 37%LMArena2026-10-08

Writing & Preference

GPT-6 Sol Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1395#124 of 297, top 42%LMArena2026-10-08
LMArena Creative Writing1378#105 of 295, top 36%LMArena2026-10-08
EQ-Bench Creative Writing2125#4 of 115, top 4%EQ-Bench
LMArena Multi-Turn1412#105 of 295, top 36%LMArena2026-10-08

API pricing by provider

GPT-6 Sol API prices
RouteInput $/MOutput $/MCached input $/MChecked
azure$2$10$0.202026-10-10
bedrock$2$10$0.202026-10-10
openai$2$10$0.202026-10-10
openrouter$2$10$0.202026-10-10

Compare GPT-6 Sol

Other OpenAI models

Frequently asked questions

How good is GPT-6 Sol?

GPT-6 Sol by OpenAI ranks 12th of 354 ranked models on the Noometry Index as of October 2026, with a score of 61.8. Its strongest category is math, where it ranks 7th. API pricing starts at $2 per million input tokens and $10 per million output tokens, with a 1.05M-token context window.

How much does GPT-6 Sol cost?

GPT-6 Sol costs $2 per million input tokens and $10 per million output tokens on OpenAI's own API, with cached input at $0.20.

What is GPT-6 Sol's context window?

GPT-6 Sol accepts up to 1.05M tokens of input and can write up to 128K tokens in one response.

Is GPT-6 Sol open source?

No. GPT-6 Sol is proprietary and available only through OpenAI's API and partner platforms.

What are GPT-6 Sol's strengths and weaknesses?

Relative to other ranked models, GPT-6 Sol places best in math, reasoning, coding and lowest in multilingual, long context, instruction following.

What is GPT-6 Sol best at?

Its best category is math, where it ranks 7th on Noometry.