Model comparison

GLM-4.6 vs Nemotron 3.5 Lightning

GLM-4.6 is the stronger model overall, scoring 41.4 to 40.0 on the Noometry Index. Nemotron 3.5 Lightning costs 11× less per token, which makes it the better buy when GLM-4.6's lead doesn't matter for your workload.

Last verified . 18 shared benchmarks.

GLM-4.6 Z.ai (Zhipu)

41.4

Rank #135 Confirmed

Nemotron 3.5 Lightning NVIDIA

40.0

Rank #155 Confirmed

Summary

  • They share 18 benchmarks with published results for both. GLM-4.6 scores higher in 6 categories and Nemotron 3.5 Lightning in 2 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where GLM-4.6 leads 61.1 to 48.5.
  • Nemotron 3.5 Lightning is cheaper at $0.05 / $0.20 per million input/output tokens, against $0.60 / $2.20 for GLM-4.6.
  • Nemotron 3.5 Lightning accepts more context: 262K tokens versus 205K.

Side by side

GLM-4.6 and Nemotron 3.5 Lightning specifications
GLM-4.6Nemotron 3.5 Lightning
ProviderZ.ai (Zhipu)NVIDIA
Noometry Index41.440.0
Released2025-09-302026-08-11
WeightsOpenOpen
Context window205K262K
Max output131K262K
Input $ / M tokens$0.60$0.05
Output $ / M tokens$2.20$0.20
Results tracked2918

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

GLM-4.6: 40.1 (#148), Nemotron 3.5 Lightning: 40.4 (#141)

Coding benchmarks
BenchmarkGLM-4.6Nemotron 3.5 Lightning
LMArena Coding14491375
SWE-bench Verified (bash only)55.4%—
LMArena WebDev1340—
SciCode38.4%—
ALE-Bench340.82—

Agentic & Tool Use Not comparable

GLM-4.6: 32.3 (#66), Nemotron 3.5 Lightning: —

Agentic & Tool Use benchmarks
BenchmarkGLM-4.6Nemotron 3.5 Lightning
Terminal-Bench24.5%—
Berkeley Function Calling Leaderboard72.4%—

Reasoning Nemotron 3.5 Lightning leads

GLM-4.6: 23.7 (#172), Nemotron 3.5 Lightning: 26.8 (#127)

Reasoning benchmarks
BenchmarkGLM-4.6Nemotron 3.5 Lightning
LMArena Hard Prompts14401337
Kagi LLM Benchmark47.4%—
CritPt1.1%—

Math GLM-4.6 leads

GLM-4.6: 39.1 (#111), Nemotron 3.5 Lightning: 37.5 (#155)

Math benchmarks
BenchmarkGLM-4.6Nemotron 3.5 Lightning
LMArena Math14321359
FrontierMath (Feb 2025 set)3.8%—
FrontierMath Tier 4 (v1)2.1%—

Knowledge GLM-4.6 leads

GLM-4.6: 40.2 (#124), Nemotron 3.5 Lightning: 37.5 (#154)

Knowledge benchmarks
BenchmarkGLM-4.6Nemotron 3.5 Lightning
LMArena Expert14311356
Vectara Hallucination Rate9.5%—

Multilingual GLM-4.6 leads

GLM-4.6: 53.5 (#66), Nemotron 3.5 Lightning: 44.0 (#180)

Multilingual benchmarks
BenchmarkGLM-4.6Nemotron 3.5 Lightning
LMArena Non-English14261295
LMArena Chinese14991359
LMArena French14591366
LMArena German14471282
LMArena Japanese13931206
LMArena Korean14001238
LMArena Russian14191253
LMArena Spanish14361345

Instruction Following GLM-4.6 leads

GLM-4.6: 74.3 (#98), Nemotron 3.5 Lightning: 69.6 (#170)

Instruction Following benchmarks
BenchmarkGLM-4.6Nemotron 3.5 Lightning
LMArena Instruction Following14101318

Long Context GLM-4.6 leads

GLM-4.6: 43.4 (#94), Nemotron 3.5 Lightning: 39.9 (#165)

Long Context benchmarks
BenchmarkGLM-4.6Nemotron 3.5 Lightning
LMArena Longer Query14221314

Writing & Preference GLM-4.6 leads

GLM-4.6: 61.1 (#90), Nemotron 3.5 Lightning: 48.5 (#201)

Writing & Preference benchmarks
BenchmarkGLM-4.6Nemotron 3.5 Lightning
LMArena Text14401327
LMArena Creative Writing14111254
EQ-Bench Creative Writing14111280
LMArena Multi-Turn14271328

Frequently asked questions

Is GLM-4.6 better than Nemotron 3.5 Lightning?

GLM-4.6 is the stronger model overall, scoring 41.4 to 40.0 on the Noometry Index. Nemotron 3.5 Lightning costs 11× less per token, which makes it the better buy when GLM-4.6's lead doesn't matter for your workload.

Which is cheaper, GLM-4.6 or Nemotron 3.5 Lightning?

Nemotron 3.5 Lightning is cheaper. It lists at $0.05 per million input tokens and $0.20 per million output tokens; GLM-4.6 lists at $0.60 and $2.20.

Is GLM-4.6 or Nemotron 3.5 Lightning better for coding?

They score almost the same on coding (40.1 vs 40.4); test both on your own repository before choosing.

Which has the bigger context window?

Nemotron 3.5 Lightning does, with 262K tokens against 205K.

How many benchmarks do GLM-4.6 and Nemotron 3.5 Lightning share?

18 benchmarks have published results for both models. GLM-4.6 has 29 scored results on Noometry and Nemotron 3.5 Lightning has 18.

Related comparisons

Go deeper