Model comparison

Amazon Nova Experimental Chat 10 09 vs GLM-4.5

Amazon Nova Experimental Chat 10 09 and GLM-4.5 score almost the same on the Noometry Index (41.9 vs 42.0), so choose on price, context window or the category you care about most.

Last verified . 9 shared benchmarks.

GLM-4.5 Z.ai (Zhipu)

42.0

Rank #122 Confirmed

Summary

  • They share 9 benchmarks with published results for both. Amazon Nova Experimental Chat 10 09 scores higher in 1 category and GLM-4.5 in 5 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in multilingual, where GLM-4.5 leads 52.8 to 47.3.
  • GLM-4.5 has downloadable open weights; the other is API-only.

Side by side

Amazon Nova Experimental Chat 10 09 and GLM-4.5 specifications
Amazon Nova Experimental Chat 10 09GLM-4.5
ProviderAmazonZ.ai (Zhipu)
Noometry Index41.942.0
Released—2025-07-27
WeightsProprietaryOpen
Context window—131K
Max output—98K
Input $ / M tokens—$0.60
Output $ / M tokens—$2.20
Results tracked927

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding GLM-4.5 leads

Amazon Nova Experimental Chat 10 09: 40.2 (#145), GLM-4.5: 41.4 (#125)

Coding benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09GLM-4.5
LMArena Coding13701434
SWE-bench Verified (bash only)—54.2%
WeirdML—40.6%
ALE-Bench—344.82
AlgoTune—1.52

Reasoning GLM-4.5 leads

Amazon Nova Experimental Chat 10 09: 27.3 (#120), GLM-4.5: 28.6 (#100)

Reasoning benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09GLM-4.5
LMArena Hard Prompts13561429
Kagi LLM Benchmark—57.9%

Math Not comparable

Amazon Nova Experimental Chat 10 09: —, GLM-4.5: 39.0 (#116)

Math benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09GLM-4.5
LMArena Math—1427

Knowledge Not comparable

Amazon Nova Experimental Chat 10 09: —, GLM-4.5: 35.9 (#179)

Knowledge benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09GLM-4.5
Humanity's Last Exam—8.3%
Confabulations—11.3%
LMArena Expert—1433

Multilingual GLM-4.5 leads

Amazon Nova Experimental Chat 10 09: 47.3 (#150), GLM-4.5: 52.8 (#77)

Multilingual benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09GLM-4.5
LMArena Non-English13401417
LMArena Chinese13411465
LMArena French—1418
LMArena German—1407
LMArena Japanese—1415
LMArena Korean—1380
LMArena Russian—1414
LMArena Spanish—1454

Instruction Following GLM-4.5 leads

Amazon Nova Experimental Chat 10 09: 69.4 (#172), GLM-4.5: 74.1 (#104)

Instruction Following benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09GLM-4.5
LMArena Instruction Following13151404

Long Context Amazon Nova Experimental Chat 10 09 leads

Amazon Nova Experimental Chat 10 09: 40.5 (#152), GLM-4.5: 38.2 (#201)

Long Context benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09GLM-4.5
LMArena Longer Query13331412
Fiction.LiveBench—58.3%

Writing & Preference GLM-4.5 leads

Amazon Nova Experimental Chat 10 09: 54.7 (#149), GLM-4.5: 57.5 (#127)

Writing & Preference benchmarks
BenchmarkAmazon Nova Experimental Chat 10 09GLM-4.5
LMArena Text13641430
LMArena Creative Writing13071395
LMArena Multi-Turn13551415
Short-Story Creative Writing—73.4%
EQ-Bench Creative Writing—1343

Frequently asked questions

Is Amazon Nova Experimental Chat 10 09 better than GLM-4.5?

Amazon Nova Experimental Chat 10 09 and GLM-4.5 score almost the same on the Noometry Index (41.9 vs 42.0), so choose on price, context window or the category you care about most.

Is Amazon Nova Experimental Chat 10 09 or GLM-4.5 better for coding?

GLM-4.5 scores higher on coding benchmarks: 41.4 versus 40.2 in the Noometry coding category.

How many benchmarks do Amazon Nova Experimental Chat 10 09 and GLM-4.5 share?

9 benchmarks have published results for both models. Amazon Nova Experimental Chat 10 09 has 9 scored results on Noometry and GLM-4.5 has 27.

Related comparisons

Go deeper