Model comparison

Amazon Nova Lite vs Grok 4.5

Grok 4.5 is the stronger model overall, scoring 55.0 to 31.9 on the Noometry Index. Amazon Nova Lite costs 29× less per token, which makes it the better buy when Grok 4.5's lead doesn't matter for your workload.

Last verified . 19 shared benchmarks.

Amazon Nova Lite Amazon

31.9

Rank #265 Confirmed

Grok 4.5 xAI

55.0

Rank #25 Confirmed

Summary

  • They share 19 benchmarks with published results for both. Amazon Nova Lite scores higher in 0 categories and Grok 4.5 in 9 categories; 9 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Grok 4.5 leads 56.1 to 19.3.
  • Amazon Nova Lite is cheaper at $0.06 / $0.24 per million input/output tokens, against $2 / $6 for Grok 4.5.
  • Grok 4.5 accepts more context: 500K tokens versus 300K.

Side by side

Amazon Nova Lite and Grok 4.5 specifications
Amazon Nova LiteGrok 4.5
ProviderAmazonxAI
Noometry Index31.955.0
Released2024-12-032026-07-08
WeightsProprietaryProprietary
Context window300K500K
Max output10K500K
Input $ / M tokens$0.06$2
Output $ / M tokens$0.24$6
Results tracked3452

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Grok 4.5 leads

Amazon Nova Lite: 32.5 (#270), Grok 4.5: 52.2 (#35)

Coding benchmarks
BenchmarkAmazon Nova LiteGrok 4.5
LMArena Coding12391474
ALE-Bench236.251,309
DeepSWE—53.8%
FrontierCode—42.4%
LMArena WebDev—1553
SciCode—54.1%
WeirdML—46.4%
LiveBench Coding27.5%—

Agentic & Tool Use Not comparable

Amazon Nova Lite: —, Grok 4.5: 44.4 (#17)

Agentic & Tool Use benchmarks
BenchmarkAmazon Nova LiteGrok 4.5
APEX-Agents—56.2%
τ²-bench Banking—47.9%
PostTrainBench—23.4%
GBAEval—65.4%
GDP.pdf—14%
LMArena Search—1213
Vending-Bench 2—3,887

Reasoning Grok 4.5 leads

Amazon Nova Lite: 19.3 (#260), Grok 4.5: 56.1 (#25)

Reasoning benchmarks
BenchmarkAmazon Nova LiteGrok 4.5
LMArena Hard Prompts12201462
ARC-AGI-2—52.6%
SimpleBench—70%
Kagi LLM Benchmark—83.5%
NYT Connections (extended)—79.9%
ARC-AGI-1—87.2%
CritPt—15.4%
Chess Puzzles—36%
LiveBench Reasoning36.7%—
DTBench—96.5%
LiveBench Data Analysis37.2%—
LMCA—45.2%
Surface Evolver Bench—74.4%
Epoch Capabilities Index—153.92
LiveBench36.4%—

Math Grok 4.5 leads

Amazon Nova Lite: 27.8 (#247), Grok 4.5: 60.9 (#35)

Math benchmarks
BenchmarkAmazon Nova LiteGrok 4.5
LMArena Math12271459
FrontierMath (Tiers 1-3)—57.2%
FrontierMath Tier 4—24.4%
OTIS Mock AIME 2024-2025—97.8%
ProofBench—31%
Omni-MATH23.3%—
LiveBench Math36.7%—

Knowledge Grok 4.5 leads

Amazon Nova Lite: 26.3 (#259), Grok 4.5: 62.3 (#24)

Knowledge benchmarks
BenchmarkAmazon Nova LiteGrok 4.5
LMArena Expert12011466
GPQA Diamond—93.4%
Humanity's Last Exam3.6%—
SimpleQA Verified—48.3%
MMLU-Pro60%—
Vectara Hallucination Rate6.1%—
GPQA (HELM)39.7%—
MMLU77%—

Multimodal Grok 4.5 leads

Amazon Nova Lite: 25.5 (#123), Grok 4.5: 37.6 (#72)

Multimodal benchmarks
BenchmarkAmazon Nova LiteGrok 4.5
LMArena Vision9901288
Blueprint-Bench 2—27.3%
Furniture Assembly—22.5%
LMArena Document—1452

Multilingual Grok 4.5 leads

Amazon Nova Lite: 38.0 (#232), Grok 4.5: 54.4 (#42)

Multilingual benchmarks
BenchmarkAmazon Nova LiteGrok 4.5
LMArena Non-English12081440
LMArena Chinese12251496
LMArena French12371456
LMArena German12291446
LMArena Japanese11531428
LMArena Korean11531404
LMArena Russian12161448
LMArena Spanish12261450

Instruction Following Grok 4.5 leads

Amazon Nova Lite: 59.3 (#255), Grok 4.5: 76.0 (#48)

Instruction Following benchmarks
BenchmarkAmazon Nova LiteGrok 4.5
LMArena Instruction Following12051446
LiveBench Instruction Following54.1%—
IFEval77.6%—

Long Context Grok 4.5 leads

Amazon Nova Lite: 37.4 (#217), Grok 4.5: 44.8 (#56)

Long Context benchmarks
BenchmarkAmazon Nova LiteGrok 4.5
LMArena Longer Query12341463

Writing & Preference Grok 4.5 leads

Amazon Nova Lite: 42.1 (#237), Grok 4.5: 65.8 (#42)

Writing & Preference benchmarks
BenchmarkAmazon Nova LiteGrok 4.5
LMArena Text12291448
LMArena Creative Writing11971442
LMArena Multi-Turn11991456
EQ-Bench Creative Writing—1579
WildBench75%—
LiveBench Language25.9%—

Frequently asked questions

Is Amazon Nova Lite better than Grok 4.5?

Grok 4.5 is the stronger model overall, scoring 55.0 to 31.9 on the Noometry Index. Amazon Nova Lite costs 29× less per token, which makes it the better buy when Grok 4.5's lead doesn't matter for your workload.

Which is cheaper, Amazon Nova Lite or Grok 4.5?

Amazon Nova Lite is cheaper. It lists at $0.06 per million input tokens and $0.24 per million output tokens; Grok 4.5 lists at $2 and $6.

Is Amazon Nova Lite or Grok 4.5 better for coding?

Grok 4.5 scores higher on coding benchmarks: 52.2 versus 32.5 in the Noometry coding category.

Which has the bigger context window?

Grok 4.5 does, with 500K tokens against 300K.

How many benchmarks do Amazon Nova Lite and Grok 4.5 share?

19 benchmarks have published results for both models. Amazon Nova Lite has 34 scored results on Noometry and Grok 4.5 has 52.

Related comparisons

Go deeper