Model comparison

GPT-5 Nano vs Nova 2.0 Pro Preview

GPT-5 Nano and Nova 2.0 Pro Preview score almost the same on the Noometry Index (33.5 vs 33.4), so choose on price, context window or the category you care about most.

Last verified . 0 shared benchmarks.

GPT-5 Nano OpenAI

33.5

Rank #241 Confirmed

Nova 2.0 Pro Preview Amazon

33.4

Rank #244 Reported

Summary

  • The widest gap is in coding, where Nova 2.0 Pro Preview leads 40.8 to 33.6.

Side by side

GPT-5 Nano and Nova 2.0 Pro Preview specifications
GPT-5 NanoNova 2.0 Pro Preview
ProviderOpenAIAmazon
Noometry Index33.533.4
Released2025-08-072025-12-02
WeightsProprietaryProprietary
Context window400K—
Max output128K—
Input $ / M tokens$0.05—
Output $ / M tokens$0.40—
Results tracked493

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Nova 2.0 Pro Preview leads

GPT-5 Nano: 33.6 (#254), Nova 2.0 Pro Preview: 40.8 (#133)

Coding benchmarks
BenchmarkGPT-5 NanoNova 2.0 Pro Preview
SWE-bench Verified (bash only)34.8%—
SciCode—42.7%
WeirdML38.1%—
LMArena Coding1351—
ALE-Bench718.67—

Agentic & Tool Use GPT-5 Nano leads

GPT-5 Nano: 25.8 (#106), Nova 2.0 Pro Preview: 21.9 (#136)

Agentic & Tool Use benchmarks
BenchmarkGPT-5 NanoNova 2.0 Pro Preview
Terminal-Bench21.8%—
Berkeley Function Calling Leaderboard51.5%—
GDP.pdf—2%

Reasoning Nova 2.0 Pro Preview leads

GPT-5 Nano: 16.3 (#306), Nova 2.0 Pro Preview: 22.4 (#194)

Reasoning benchmarks
BenchmarkGPT-5 NanoNova 2.0 Pro Preview
ARC-AGI-22.6%—
Kagi LLM Benchmark62.2%—
ARC-AGI-120.7%—
CritPt—0%
Chess Puzzles27%—
LMArena Hard Prompts1328—
Mystery Game Puzzles9%—
DTBench62.7%—
LMCA7.9%—
Epoch Capabilities Index139.38—
ForecastBench59.1—

Math Not comparable

GPT-5 Nano: 29.4 (#241), Nova 2.0 Pro Preview: —

Math benchmarks
BenchmarkGPT-5 NanoNova 2.0 Pro Preview
FrontierMath (Tiers 1-3)20%—
FrontierMath Tier 42.4%—
OTIS Mock AIME 2024-202581.1%—
ProofBench12%—
Omni-MATH54.6%—
LMArena Math1317—
MATH Level 595.2%—
FrontierMath (Feb 2025 set)8.3%—
FrontierMath Tier 4 (v1)2.1%—

Knowledge Not comparable

GPT-5 Nano: 35.9 (#178), Nova 2.0 Pro Preview: —

Knowledge benchmarks
BenchmarkGPT-5 NanoNova 2.0 Pro Preview
GPQA Diamond69.4%—
SimpleQA Verified11.7%—
MMLU-Pro77.8%—
Vectara Hallucination Rate10.5%—
GPQA (HELM)67.9%—
LMArena Expert1321—

Multimodal Not comparable

GPT-5 Nano: 31.3 (#108), Nova 2.0 Pro Preview: —

Multimodal benchmarks
BenchmarkGPT-5 NanoNova 2.0 Pro Preview
LMArena Vision1159—
VPCT37.2%—

Multilingual Not comparable

GPT-5 Nano: 45.3 (#172), Nova 2.0 Pro Preview: —

Multilingual benchmarks
BenchmarkGPT-5 NanoNova 2.0 Pro Preview
LMArena Non-English1313—
LMArena Chinese1356—
LMArena German1327—
LMArena Japanese1226—
LMArena Korean1269—
LMArena Russian1296—
LMArena Spanish1360—

Instruction Following Not comparable

GPT-5 Nano: 75.0 (#79), Nova 2.0 Pro Preview: —

Instruction Following benchmarks
BenchmarkGPT-5 NanoNova 2.0 Pro Preview
IFEval93.2%—
LMArena Instruction Following1306—

Long Context Not comparable

GPT-5 Nano: 31.3 (#281), Nova 2.0 Pro Preview: —

Long Context benchmarks
BenchmarkGPT-5 NanoNova 2.0 Pro Preview
Fiction.LiveBench44.4%—
LMArena Longer Query1312—

Writing & Preference Not comparable

GPT-5 Nano: 39.1 (#249), Nova 2.0 Pro Preview: —

Writing & Preference benchmarks
BenchmarkGPT-5 NanoNova 2.0 Pro Preview
LMArena Text1320—
LMArena Creative Writing1249—
EQ-Bench Creative Writing705—
WildBench80.6%—
LMArena Multi-Turn1311—

Frequently asked questions

Is GPT-5 Nano better than Nova 2.0 Pro Preview?

GPT-5 Nano and Nova 2.0 Pro Preview score almost the same on the Noometry Index (33.5 vs 33.4), so choose on price, context window or the category you care about most.

Is GPT-5 Nano or Nova 2.0 Pro Preview better for coding?

Nova 2.0 Pro Preview scores higher on coding benchmarks: 40.8 versus 33.6 in the Noometry coding category.

How many benchmarks do GPT-5 Nano and Nova 2.0 Pro Preview share?

0 benchmarks have published results for both models. GPT-5 Nano has 49 scored results on Noometry and Nova 2.0 Pro Preview has 3.

Related comparisons

Go deeper