Model comparison

Amazon Nova Experimental Chat 10 20 vs GLM-4.5

Amazon Nova Experimental Chat 10 20 and GLM-4.5 score almost the same on the Noometry Index (42.1 vs 42.0), so choose on price, context window or the category you care about most.

Last verified . 17 shared benchmarks.

GLM-4.5 Z.ai (Zhipu)

42.0

Rank #122 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Amazon Nova Experimental Chat 10 20 scores higher in 3 categories and GLM-4.5 in 5 categories; 4 gaps are clear of the uncertainty.
  • The widest gap is in long context, where Amazon Nova Experimental Chat 10 20 leads 41.7 to 38.2.
  • GLM-4.5 has downloadable open weights; the other is API-only.

Side by side

Amazon Nova Experimental Chat 10 20 and GLM-4.5 specifications
Amazon Nova Experimental Chat 10 20GLM-4.5
ProviderAmazonZ.ai (Zhipu)
Noometry Index42.142.0
Released—2025-07-27
WeightsProprietaryOpen
Context window—131K
Max output—98K
Input $ / M tokens—$0.60
Output $ / M tokens—$2.20
Results tracked1727

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Amazon Nova Experimental Chat 10 20: 41.6 (#123), GLM-4.5: 41.4 (#125)

Coding benchmarks
BenchmarkAmazon Nova Experimental Chat 10 20GLM-4.5
LMArena Coding14121434
SWE-bench Verified (bash only)—54.2%
WeirdML—40.6%
ALE-Bench—344.82
AlgoTune—1.52

Reasoning Too close to call

Amazon Nova Experimental Chat 10 20: 28.4 (#106), GLM-4.5: 28.6 (#100)

Reasoning benchmarks
BenchmarkAmazon Nova Experimental Chat 10 20GLM-4.5
LMArena Hard Prompts13961429
Kagi LLM Benchmark—57.9%

Math Too close to call

Amazon Nova Experimental Chat 10 20: 39.0 (#119), GLM-4.5: 39.0 (#116)

Math benchmarks
BenchmarkAmazon Nova Experimental Chat 10 20GLM-4.5
LMArena Math14251427

Knowledge Amazon Nova Experimental Chat 10 20 leads

Amazon Nova Experimental Chat 10 20: 38.5 (#144), GLM-4.5: 35.9 (#179)

Knowledge benchmarks
BenchmarkAmazon Nova Experimental Chat 10 20GLM-4.5
LMArena Expert13861433
Humanity's Last Exam—8.3%
Confabulations—11.3%

Multilingual GLM-4.5 leads

Amazon Nova Experimental Chat 10 20: 49.5 (#131), GLM-4.5: 52.8 (#77)

Multilingual benchmarks
BenchmarkAmazon Nova Experimental Chat 10 20GLM-4.5
LMArena Non-English13721417
LMArena Chinese14041465
LMArena French14251418
LMArena German13911407
LMArena Japanese13601415
LMArena Korean13241380
LMArena Russian13711414
LMArena Spanish13821454

Instruction Following GLM-4.5 leads

Amazon Nova Experimental Chat 10 20: 72.1 (#141), GLM-4.5: 74.1 (#104)

Instruction Following benchmarks
BenchmarkAmazon Nova Experimental Chat 10 20GLM-4.5
LMArena Instruction Following13651404

Long Context Amazon Nova Experimental Chat 10 20 leads

Amazon Nova Experimental Chat 10 20: 41.7 (#133), GLM-4.5: 38.2 (#201)

Long Context benchmarks
BenchmarkAmazon Nova Experimental Chat 10 20GLM-4.5
LMArena Longer Query13701412
Fiction.LiveBench—58.3%

Writing & Preference Too close to call

Amazon Nova Experimental Chat 10 20: 56.6 (#136), GLM-4.5: 57.5 (#127)

Writing & Preference benchmarks
BenchmarkAmazon Nova Experimental Chat 10 20GLM-4.5
LMArena Text13941430
LMArena Creative Writing13181395
LMArena Multi-Turn13641415
Short-Story Creative Writing—73.4%
EQ-Bench Creative Writing—1343

Frequently asked questions

Is Amazon Nova Experimental Chat 10 20 better than GLM-4.5?

Amazon Nova Experimental Chat 10 20 and GLM-4.5 score almost the same on the Noometry Index (42.1 vs 42.0), so choose on price, context window or the category you care about most.

Is Amazon Nova Experimental Chat 10 20 or GLM-4.5 better for coding?

They score almost the same on coding (41.6 vs 41.4); test both on your own repository before choosing.

How many benchmarks do Amazon Nova Experimental Chat 10 20 and GLM-4.5 share?

17 benchmarks have published results for both models. Amazon Nova Experimental Chat 10 20 has 17 scored results on Noometry and GLM-4.5 has 27.

Related comparisons

Go deeper