OpenAI, proprietary
GPT-5.4 nano
GPT-5.4 nano by OpenAI ranks 125th of 354 ranked models on the Noometry Index as of October 2026, with a score of 41.9. Its strongest category is multimodal, where it ranks 78th. API pricing starts at $0.20 per million input tokens and $1.25 per million output tokens, with a 400K-token context window.
Last verified
Specifications
- Noometry rank
- #125 of 354
- Index score
- 41.9
- Evidence
- Confirmed 40 results
- Provider
- OpenAI
- Released
- March 17, 2026
- Weights
- Proprietary
- Reasoning
- Yes
- Context window
- 400K
- Max output
- 128K
- Input price
- $0.20 / M
- Output price
- $1.25 / M
- Blended price
- $0.46 / M
- Output speed
- 19 tokens/s Kagi
- Value
- #71 of 219
- Knowledge cutoff
- August 2025
- Input
- text, image
Category scores
Each category score combines every public result we have in that category.
- Coding 43.6
- Reasoning 23.7
- Math 40.9
- Knowledge 41.9
- Multimodal 36.7
- Multilingual 48.6
- Instruction Following 71.9
- Long Context 41.6
- Writing & Preference 55.7
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 43.6 | #84 | 3 |
| Reasoning | 23.7 | #173 | 9 |
| Math | 40.9 | #88 | 5 |
| Knowledge | 41.9 | #103 | 4 |
| Multimodal | 36.7 | #78 | 1 |
| Multilingual | 48.6 | #140 | 1 |
| Instruction Following | 71.9 | #144 | 1 |
| Long Context | 41.6 | #137 | 1 |
| Writing & Preference | 55.7 | #142 | 3 |
Strengths and weaknesses
Categories where GPT-5.4 nano places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Multimodal | 36.7 | −1.8 | #78 of 128, top 61% |
| Reasoning | 23.7 | +0.1 | #173 of 350, top 50% |
| Instruction Following | 71.9 | +0.7 | #144 of 305, top 48% |
Closest competitors
The models ranked just above and below GPT-5.4 nano. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Solar Pro4 | #121 | 42.1 | $0.52 | — | Compare |
| GLM-4.5 | #122 | 42.0 | $1 | 32 | Compare |
| Qwen3.5 35B-A3B | #123 | 42.0 | $0.69 | — | Compare |
| GLM-4.7 | #124 | 42.0 | $1 | — | Compare |
| Amazon Nova Experimental Chat 10 09 | #126 | 41.9 | — | — | Compare |
| Qwen3.5 27B | #127 | 41.9 | $0.82 | — | Compare |
| GPT-5 Mini | #128 | 41.8 | $0.69 | 3 | Compare |
| ERNIE 5.0 0110 | #129 | 41.8 | — | — | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| SciCode | 46.9% | #55 of 121, top 46% | xhigh | Epoch AI | |
| WeirdML | 49.2% | #49 of 119, top 42% | high | Epoch AI | |
| WeirdML | 38% | none | Epoch AI | ||
| LMArena Coding | 1405 | #134 of 294, top 46% | high | LMArena | 2026-10-08 |
| ALE-Bench | 1,005 | #42 of 105, top 40% | high | Epoch AI |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| ARC-AGI-2 | 3.6% | high | Epoch AI | ||
| ARC-AGI-2 | 1.5% | low | Epoch AI | ||
| ARC-AGI-2 | 1.9% | medium | Epoch AI | ||
| ARC-AGI-2 | 5.7% | #54 of 83, top 66% | xhigh | Epoch AI | |
| Kagi LLM Benchmark | 39.7% | #80 of 99, top 81% | Kagi LLM Benchmark | ||
| ARC-AGI-1 | 38.2% | high | Epoch AI | ||
| ARC-AGI-1 | 18.3% | low | Epoch AI | ||
| ARC-AGI-1 | 33% | medium | Epoch AI | ||
| ARC-AGI-1 | 51.5% | #56 of 83, top 68% | xhigh | Epoch AI | |
| CritPt | 9.3% | #49 of 134, top 37% | xhigh | Epoch AI | |
| Chess Puzzles | 30% | #36 of 129, top 28% | high | Epoch AI | 2026-04-14 |
| Chess Puzzles | 17% | low | Epoch AI | 2026-08-07 | |
| Chess Puzzles | 3% | none | Epoch AI | 2026-08-07 | |
| LMArena Hard Prompts | 1381 | #137 of 297, top 47% | high | LMArena | 2026-10-08 |
| Mystery Game Puzzles | 5% | high | Epoch AI | 2026-08-27 | |
| Mystery Game Puzzles | 3% | low | Epoch AI | 2026-08-27 | |
| Mystery Game Puzzles | 6% | medium | Epoch AI | 2026-08-27 | |
| Mystery Game Puzzles | 9% | #61 of 74, top 83% | none | Epoch AI | 2026-08-27 |
| DTBench | 80.3% | #73 of 151, top 49% | xhigh | Epoch AI | |
| LMCA | 36.9% | #60 of 125, top 48% | xhigh | Epoch AI | |
| Epoch Capabilities Index | 145.81 | #79 of 213, top 38% | Epoch AI | 2026-03-17 | |
| ForecastBench | 57.3 | #59 of 72, top 82% | Epoch AI |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| FrontierMath (Tiers 1-3) | 44.9% | #50 of 81, top 62% | high | Epoch AI | 2026-06-12 |
| FrontierMath (Tiers 1-3) | 20.4% | low | Epoch AI | 2026-08-28 | |
| FrontierMath (Tiers 1-3) | 4.6% | none | Epoch AI | 2026-08-28 | |
| FrontierMath Tier 4 | 12.2% | #50 of 63, top 80% | high | Epoch AI | 2026-06-12 |
| OTIS Mock AIME 2024-2025 | 87.8% | #60 of 173, top 35% | high | Epoch AI | 2026-04-14 |
| OTIS Mock AIME 2024-2025 | 68.9% | low | Epoch AI | 2026-08-07 | |
| OTIS Mock AIME 2024-2025 | 46.7% | none | Epoch AI | 2026-08-07 | |
| ProofBench | 5% | #69 of 77, top 90% | high | Epoch AI | |
| LMArena Math | 1406 | #120 of 285, top 43% | high | LMArena | 2026-10-08 |
| FrontierMath (Feb 2025 set) | 25.9% | #24 of 68, top 36% | high | Epoch AI | 2026-04-15 |
| FrontierMath Tier 4 (v1) | 6.3% | #23 of 55, top 42% | high | Epoch AI | 2026-04-15 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| GPQA Diamond | 78.5% | #84 of 186, top 46% | high | Epoch AI | 2026-04-14 |
| GPQA Diamond | 72.2% | low | Epoch AI | 2026-08-07 | |
| GPQA Diamond | 55.6% | none | Epoch AI | 2026-08-07 | |
| SimpleQA Verified | 11.7% | #73 of 77, top 95% | high | Epoch AI | 2026-08-27 |
| Vectara Hallucination Rate (lower is better) | 3.1% | Best of 96 | Vectara Hallucination Leaderboard | ||
| LMArena Expert | 1396 | #126 of 273, top 47% | high | LMArena | 2026-10-08 |
Multimodal
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Vision | 1196 | #80 of 122, top 66% | high | LMArena | 2026-10-09 |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1359 | #140 of 297, top 48% | high | LMArena | 2026-10-08 |
| LMArena Chinese | 1392 | #142 of 285, top 50% | high | LMArena | 2026-10-08 |
| LMArena French | 1396 | #123 of 223, top 56% | high | LMArena | 2026-10-08 |
| LMArena German | 1367 | #118 of 231, top 52% | high | LMArena | 2026-10-08 |
| LMArena Japanese | 1343 | #102 of 211, top 49% | high | LMArena | 2026-10-08 |
| LMArena Korean | 1320 | #119 of 213, top 56% | high | LMArena | 2026-10-08 |
| LMArena Russian | 1363 | #139 of 283, top 50% | high | LMArena | 2026-10-08 |
| LMArena Spanish | 1371 | #133 of 226, top 59% | high | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Instruction Following | 1362 | #138 of 298, top 47% | high | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 1366 | #144 of 291, top 50% | high | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1372 | #143 of 297, top 49% | high | LMArena | 2026-10-08 |
| LMArena Creative Writing | 1314 | #160 of 295, top 55% | high | LMArena | 2026-10-08 |
| LMArena Multi-Turn | 1382 | #133 of 295, top 46% | high | LMArena | 2026-10-08 |
API pricing by provider
| Route | Input $/M | Output $/M | Cached input $/M | Checked |
|---|---|---|---|---|
| azure | $0.20 | $1.25 | $0.02 | 2026-10-10 |
| openai | $0.20 | $1.25 | $0.02 | 2026-10-10 |
| openrouter | $0.20 | $1.25 | $0.02 | 2026-10-10 |
Compare GPT-5.4 nano
- GPT-5.4 nano vs GPT-5 Nano
- GPT-5.4 nano vs GLM-4.7
- GPT-5.4 nano vs Amazon Nova Experimental Chat 10 09
- GPT-5.4 nano vs Qwen3.5 35B-A3B
- GPT-5.4 nano vs Qwen3.5 27B
- GPT-5.4 nano vs GLM-4.5
- GPT-5.4 nano vs GPT-5 Mini
- GPT-5.4 nano vs Claude Fable 5.1
- GPT-5.4 nano vs Gemini 3.8 Flash
- GPT-5.4 nano vs Kimi K3
- GPT-5.4 nano vs Grok 4.6
- GPT-5.4 nano vs Qwen3.8 Max
- GPT-5.4 nano vs GLM-5.3
- GPT-5.4 nano vs Muse Spark 1.3
Other OpenAI models
- GPT-6 Astra70.8
- GPT-6.1 Sol65.6
- GPT-5.6 Sol65.0
- GPT-5.5 Pro64.3
- GPT-5.563.4
- GPT-6 Sol61.8
- GPT-5.459.4
- GPT-5.6 Terra59.2
Frequently asked questions
How good is GPT-5.4 nano?
GPT-5.4 nano by OpenAI ranks 125th of 354 ranked models on the Noometry Index as of October 2026, with a score of 41.9. Its strongest category is multimodal, where it ranks 78th. API pricing starts at $0.20 per million input tokens and $1.25 per million output tokens, with a 400K-token context window.
How much does GPT-5.4 nano cost?
GPT-5.4 nano costs $0.20 per million input tokens and $1.25 per million output tokens on OpenAI's own API, with cached input at $0.02.
What is GPT-5.4 nano's context window?
GPT-5.4 nano accepts up to 400K tokens of input and can write up to 128K tokens in one response.
Is GPT-5.4 nano open source?
No. GPT-5.4 nano is proprietary and available only through OpenAI's API and partner platforms.
How fast is GPT-5.4 nano?
GPT-5.4 nano generated about 19 output tokens per second in the Kagi LLM Benchmark's timed runs. Speed varies by provider, load and reasoning effort.
What are GPT-5.4 nano's strengths and weaknesses?
Relative to other ranked models, GPT-5.4 nano places best in coding, math, knowledge and lowest in multimodal, reasoning, instruction following.
What is GPT-5.4 nano best at?
Its best category is multimodal, where it ranks 78th on Noometry.