Model comparison

Amazon Nova Pro vs GPT-6.1 Sol

GPT-6.1 Sol is the stronger model overall, scoring 65.6 to 31.0 on the Noometry Index. Amazon Nova Pro costs 2.9× less per token, which makes it the better buy when GPT-6.1 Sol's lead doesn't matter for your workload.

Last verified . 14 shared benchmarks.

Amazon Nova Pro Amazon

31.0

Rank #281 Confirmed

GPT-6.1 Sol OpenAI

65.6

Rank #6 Confirmed

Summary

  • They share 14 benchmarks with published results for both. Amazon Nova Pro scores higher in 0 categories and GPT-6.1 Sol in 10 categories; 10 gaps are clear of the uncertainty.
  • The widest gap is in math, where GPT-6.1 Sol leads 93.7 to 28.5.
  • Amazon Nova Pro is cheaper at $0.80 / $3.20 per million input/output tokens, against $2 / $10 for GPT-6.1 Sol.
  • GPT-6.1 Sol accepts more context: 1.05M tokens versus 300K.

Side by side

Amazon Nova Pro and GPT-6.1 Sol specifications
Amazon Nova ProGPT-6.1 Sol
ProviderAmazonOpenAI
Noometry Index31.065.6
Released2024-12-032026-09-29
WeightsProprietaryProprietary
Context window300K1.05M
Max output10K128K
Input $ / M tokens$0.80$2
Output $ / M tokens$3.20$10
Results tracked3834

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding GPT-6.1 Sol leads

Amazon Nova Pro: 35.1 (#229), GPT-6.1 Sol: 63.2 (#8)

Coding benchmarks
BenchmarkAmazon Nova ProGPT-6.1 Sol
LMArena Coding12701487
DeepSWE—75.2%
FrontierCode—50.2%
LMArena WebDev—1755
SciCode—55.8%
LiveBench Coding38.1%—

Agentic & Tool Use GPT-6.1 Sol leads

Amazon Nova Pro: 16.7 (#147), GPT-6.1 Sol: 39.6 (#26)

Agentic & Tool Use benchmarks
BenchmarkAmazon Nova ProGPT-6.1 Sol
APEX-Agents—60%
Berkeley Function Calling Leaderboard25%—
TheAgentCompany1.7%—
GDP.pdf—32%

Reasoning GPT-6.1 Sol leads

Amazon Nova Pro: 20.0 (#243), GPT-6.1 Sol: 81.9 (#2)

Reasoning benchmarks
BenchmarkAmazon Nova ProGPT-6.1 Sol
LMArena Hard Prompts12461466
Epoch Capabilities Index123.8166.09
ARC-AGI-2—94.2%
NYT Connections (extended)—95.5%
ARC-AGI-1—98.5%
CritPt—31.7%
Chess Puzzles—61%
EBR-Bench—54.3%
LiveBench Reasoning32.6%—
Mystery Game Puzzles—80%
LiveBench Data Analysis48.3%—
LiveBench43.5%—

Math GPT-6.1 Sol leads

Amazon Nova Pro: 28.5 (#243), GPT-6.1 Sol: 93.7 (#1)

Math benchmarks
BenchmarkAmazon Nova ProGPT-6.1 Sol
LMArena Math12521464
FrontierMath (Tiers 1-3)—93.7%
FrontierMath Tier 4—100%
OTIS Mock AIME 2024-2025—100%
ProofBench—99%
Omni-MATH24.2%—
LiveBench Math38%—

Knowledge GPT-6.1 Sol leads

Amazon Nova Pro: 27.4 (#250), GPT-6.1 Sol: 71.8 (#4)

Knowledge benchmarks
BenchmarkAmazon Nova ProGPT-6.1 Sol
LMArena Expert12111502
GPQA Diamond—95.4%
Humanity's Last Exam4.4%—
SimpleQA Verified—73.9%
MMLU-Pro67.3%—
Confabulations30.1%—
Vectara Hallucination Rate5.1%—
GPQA (HELM)44.6%—
MMLU82%—

Multimodal GPT-6.1 Sol leads

Amazon Nova Pro: 25.0 (#126), GPT-6.1 Sol: 52.7 (#5)

Multimodal benchmarks
BenchmarkAmazon Nova ProGPT-6.1 Sol
LMArena Vision9801288
Furniture Assembly—80%

Multilingual GPT-6.1 Sol leads

Amazon Nova Pro: 39.7 (#223), GPT-6.1 Sol: 54.3 (#46)

Multilingual benchmarks
BenchmarkAmazon Nova ProGPT-6.1 Sol
LMArena Non-English12341438
LMArena Chinese12441477
LMArena Russian12401455
LMArena French1271—
LMArena German1243—
LMArena Japanese1200—
LMArena Korean1203—
LMArena Spanish1182—

Instruction Following GPT-6.1 Sol leads

Amazon Nova Pro: 64.9 (#226), GPT-6.1 Sol: 77.0 (#29)

Instruction Following benchmarks
BenchmarkAmazon Nova ProGPT-6.1 Sol
LMArena Instruction Following12351468
LiveBench Instruction Following67.1%—
IFEval81.5%—

Long Context GPT-6.1 Sol leads

Amazon Nova Pro: 38.1 (#205), GPT-6.1 Sol: 44.9 (#54)

Long Context benchmarks
BenchmarkAmazon Nova ProGPT-6.1 Sol
LMArena Longer Query12551465

Writing & Preference GPT-6.1 Sol leads

Amazon Nova Pro: 43.9 (#226), GPT-6.1 Sol: 63.6 (#63)

Writing & Preference benchmarks
BenchmarkAmazon Nova ProGPT-6.1 Sol
LMArena Text12591447
LMArena Creative Writing12121432
LMArena Multi-Turn12461449
Short-Story Creative Writing60.5%—
WildBench77.7%—
LiveBench Language37%—

Frequently asked questions

Is Amazon Nova Pro better than GPT-6.1 Sol?

GPT-6.1 Sol is the stronger model overall, scoring 65.6 to 31.0 on the Noometry Index. Amazon Nova Pro costs 2.9× less per token, which makes it the better buy when GPT-6.1 Sol's lead doesn't matter for your workload.

Which is cheaper, Amazon Nova Pro or GPT-6.1 Sol?

Amazon Nova Pro is cheaper. It lists at $0.80 per million input tokens and $3.20 per million output tokens; GPT-6.1 Sol lists at $2 and $10.

Is Amazon Nova Pro or GPT-6.1 Sol better for coding?

GPT-6.1 Sol scores higher on coding benchmarks: 63.2 versus 35.1 in the Noometry coding category.

Which has the bigger context window?

GPT-6.1 Sol does, with 1.05M tokens against 300K.

How many benchmarks do Amazon Nova Pro and GPT-6.1 Sol share?

14 benchmarks have published results for both models. Amazon Nova Pro has 38 scored results on Noometry and GPT-6.1 Sol has 34.

Related comparisons

Go deeper