OpenAI, proprietary

GPT-5.3 Codex

GPT-5.3 Codex by OpenAI ranks 69th of 354 ranked models on the Noometry Index as of October 2026, with a score of 45.8. Its strongest category is agentic & tool use, where it ranks 9th. API pricing starts at $1.75 per million input tokens and $14 per million output tokens, with a 400K-token context window.

Last verified

Specifications

Noometry rank
#69 of 354
Index score
45.8
Evidence
Reported 8 results
Provider
OpenAI
Released
February 5, 2026
Weights
Proprietary
Reasoning
Yes
Context window
400K
Max output
128K
Input price
$1.75 / M
Output price
$14 / M
Blended price
$4.81 / M
Output speed
Not measured
Value
#186 of 219
Knowledge cutoff
August 2025
Input
text, image, pdf

Category scores

Each category score combines every public result we have in that category.

GPT-5.3 Codex category scores
  1. Coding 48.6
  2. Agentic & Tool Use 48.0
GPT-5.3 Codex category ranks
CategoryScoreRankResults
Coding48.6#563
Agentic & Tool Use48.0#91

Strengths and weaknesses

Categories where GPT-5.3 Codex places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

GPT-5.3 Codex: strongest categories
CategoryScorevs medianRank
Agentic & Tool Use48.0+17.7#9 of 154, top 6%

Weakest categories

GPT-5.3 Codex: weakest categories
CategoryScorevs medianRank
Coding48.6+9.9#56 of 340, top 17%

Closest competitors

The models ranked just above and below GPT-5.3 Codex. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to GPT-5.3 Codex
ModelRankScoreBlended $/MSpeed
Grok 4.20 Multi-Agent#6546.2$1.56—Compare
GLM-5#6646.1$1.5523Compare
Qwen3.5 397B-A17B#6746.0$1.359Compare
Qwen3.8 27B#6846.0$1.11—Compare
Kimi K2 Thinking Turbo#7045.8——Compare
Qwen3.5 Max Preview#7145.3——Compare
Qwen3.7 Plus#7245.3$0.70—Compare
Hy4 preview#7345.3$1.13—Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

GPT-5.3 Codex Coding benchmark results
BenchmarkScorePositionSettingSourceDate
SWE-bench Verified74.8%#15 of 32, top 47%highEpoch AI2026-02-25
LMArena WebDev1409#69 of 113, top 62%LMArena2026-10-08
WeirdML79.3%#10 of 119, top 9%Epoch AI
WeirdML77.9%xhighEpoch AI
ALE-Bench1,655#12 of 105, top 12%xhighEpoch AI

Agentic & Tool Use

GPT-5.3 Codex Agentic & Tool Use benchmark results
BenchmarkScorePositionSettingSourceDate
Terminal-Bench78.4%#6 of 41, top 15%Epoch AI
METR Time Horizons74.5%#6 of 32, top 19%Epoch AI
Vending-Bench 25,940#20 of 60, top 34%Epoch AI

Reasoning

GPT-5.3 Codex Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
Epoch Capabilities Index156.77#18 of 213, top 9%Epoch AI2026-02-05

API pricing by provider

GPT-5.3 Codex API prices
RouteInput $/MOutput $/MCached input $/MChecked
azure$1.75$14$0.172026-10-10
openai$1.75$14$0.172026-10-10
openrouter$1.75$14$0.172026-10-10

Compare GPT-5.3 Codex

Other OpenAI models

Frequently asked questions

How good is GPT-5.3 Codex?

GPT-5.3 Codex by OpenAI ranks 69th of 354 ranked models on the Noometry Index as of October 2026, with a score of 45.8. Its strongest category is agentic & tool use, where it ranks 9th. API pricing starts at $1.75 per million input tokens and $14 per million output tokens, with a 400K-token context window.

How much does GPT-5.3 Codex cost?

GPT-5.3 Codex costs $1.75 per million input tokens and $14 per million output tokens on OpenAI's own API, with cached input at $0.17.

What is GPT-5.3 Codex's context window?

GPT-5.3 Codex accepts up to 400K tokens of input and can write up to 128K tokens in one response.

Is GPT-5.3 Codex open source?

No. GPT-5.3 Codex is proprietary and available only through OpenAI's API and partner platforms.

What are GPT-5.3 Codex's strengths and weaknesses?

Relative to other ranked models, GPT-5.3 Codex places best in agentic & tool use and lowest in coding.

What is GPT-5.3 Codex best at?

Its best category is agentic & tool use, where it ranks 9th on Noometry.