Z.ai (Zhipu), open weights

GLM-4.5-Air

GLM-4.5-Air by Z.ai (Zhipu) ranks 177th of 354 ranked models on the Noometry Index as of October 2026, with a score of 38.9. Its strongest category is long context, where it ranks 135th. API pricing starts at $0.20 per million input tokens and $1.10 per million output tokens, with a 131K-token context window.

Last verified

Specifications

Noometry rank
#177 of 354
Index score
38.9
Evidence
Confirmed 27 results
Released
July 20, 2025
Weights
Open weights
Reasoning
Yes
Context window
131K
Max output
98K
Input price
$0.20 / M
Output price
$1.10 / M
Blended price
$0.43 / M
Output speed
160 tokens/s Kagi
Value
#70 of 219
Knowledge cutoff
April 2025
Input
text

Category scores

Each category score combines every public result we have in that category.

GLM-4.5-Air category scores
  1. Coding 33.3
  2. Reasoning 24.1
  3. Math 36.2
  4. Knowledge 35.0
  5. Multilingual 49.1
  6. Instruction Following 69.6
  7. Long Context 41.6
  8. Writing & Preference 55.9
GLM-4.5-Air category ranks
CategoryScoreRankResults
Coding33.3#2592
Reasoning24.1#1662
Math36.2#1702
Knowledge35.0#1915
Multilingual49.1#1351
Instruction Following69.6#1712
Long Context41.6#1351
Writing & Preference55.9#1394

Strengths and weaknesses

Categories where GLM-4.5-Air places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

GLM-4.5-Air: strongest categories
CategoryScorevs medianRank
Writing & Preference55.9+2.2#139 of 312, top 45%
Multilingual49.1+1.7#135 of 297, top 46%
Long Context41.6+0.7#135 of 296, top 46%

Weakest categories

GLM-4.5-Air: weakest categories
CategoryScorevs medianRank
Coding33.3−5.4#259 of 340, top 77%
Knowledge35.0−2.4#191 of 314, top 61%
Instruction Following69.6−1.7#171 of 305, top 57%

Closest competitors

The models ranked just above and below GLM-4.5-Air. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to GLM-4.5-Air
ModelRankScoreBlended $/MSpeed
Gemini 2.0 Pro#17339.1——Compare
Molmo 2 8b#17439.1——Compare
Mercury 2#17539.1$0.38—Compare
Mistral Large 3#17639.1$0.387Compare
MiniMax-M2.1#17838.9$0.52—Compare
Qwen3-30B-A3B#17938.9$0.2142Compare
GLM-4.7-Flash#18038.8$0.15—Compare
Qwen2.5 Plus 1127#18138.8——Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

GLM-4.5-Air Coding benchmark results
BenchmarkScorePositionSettingSourceDate
GSO2.9%#29 of 31, top 94%Epoch AI
LMArena Coding1397#138 of 294, top 47%LMArena2026-10-08

Reasoning

GLM-4.5-Air Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
Kagi LLM Benchmark43%#73 of 99, top 74%Kagi LLM Benchmark
LMArena Hard Prompts1379#139 of 297, top 47%LMArena2026-10-08
ForecastBench59.2#40 of 72, top 56%Epoch AI

Math

GLM-4.5-Air Math benchmark results
BenchmarkScorePositionSettingSourceDate
Omni-MATH39.1%#27 of 57, top 48%HELM Capabilities
LMArena Math1396#129 of 285, top 46%LMArena2026-10-08

Knowledge

GLM-4.5-Air Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
Humanity's Last Exam8.1%#27 of 41, top 66%Epoch AI
MMLU-Pro76.2%#26 of 58, top 45%HELM Capabilities
Vectara Hallucination Rate (lower is better)9.3%#44 of 96, top 46%Vectara Hallucination Leaderboard
GPQA (HELM)59.4%#24 of 57, top 43%HELM Capabilities
LMArena Expert1370#142 of 273, top 53%LMArena2026-10-08

Multilingual

GLM-4.5-Air Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1366#135 of 297, top 46%LMArena2026-10-08
LMArena Chinese1426#122 of 285, top 43%LMArena2026-10-08
LMArena French1399#118 of 223, top 53%LMArena2026-10-08
LMArena German1377#112 of 231, top 49%LMArena2026-10-08
LMArena Japanese1348#99 of 211, top 47%LMArena2026-10-08
LMArena Korean1308#126 of 213, top 60%LMArena2026-10-08
LMArena Russian1373#134 of 283, top 48%LMArena2026-10-08
LMArena Spanish1386#123 of 226, top 55%LMArena2026-10-08

Instruction Following

GLM-4.5-Air Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
IFEval81.2%#37 of 57, top 65%HELM Capabilities
LMArena Instruction Following1354#142 of 298, top 48%LMArena2026-10-08

Long Context

GLM-4.5-Air Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Longer Query1366#142 of 291, top 49%LMArena2026-10-08

Writing & Preference

GLM-4.5-Air Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1384#133 of 297, top 45%LMArena2026-10-08
LMArena Creative Writing1343#135 of 295, top 46%LMArena2026-10-08
WildBench78.9%#37 of 57, top 65%HELM Capabilities
LMArena Multi-Turn1371#138 of 295, top 47%LMArena2026-10-08

API pricing by provider

GLM-4.5-Air API prices
RouteInput $/MOutput $/MCached input $/MChecked
openrouter$0.13$0.85$0.0252026-10-10
zai$0.20$1.10$0.032026-10-10

Compare GLM-4.5-Air

Other Z.ai (Zhipu) models

Frequently asked questions

How good is GLM-4.5-Air?

GLM-4.5-Air by Z.ai (Zhipu) ranks 177th of 354 ranked models on the Noometry Index as of October 2026, with a score of 38.9. Its strongest category is long context, where it ranks 135th. API pricing starts at $0.20 per million input tokens and $1.10 per million output tokens, with a 131K-token context window.

How much does GLM-4.5-Air cost?

GLM-4.5-Air costs $0.20 per million input tokens and $1.10 per million output tokens on Z.ai (Zhipu)'s own API, with cached input at $0.03.

What is GLM-4.5-Air's context window?

GLM-4.5-Air accepts up to 131K tokens of input and can write up to 98K tokens in one response.

Is GLM-4.5-Air open source?

Yes. GLM-4.5-Air's weights are downloadable from Hugging Face (zai-org/GLM-4.5-Air); check the license for commercial terms.

How fast is GLM-4.5-Air?

GLM-4.5-Air generated about 160 output tokens per second in the Kagi LLM Benchmark's timed runs. Speed varies by provider, load and reasoning effort.

What are GLM-4.5-Air's strengths and weaknesses?

Relative to other ranked models, GLM-4.5-Air places best in writing & preference, multilingual, long context and lowest in coding, knowledge, instruction following.

What is GLM-4.5-Air best at?

Its best category is long context, where it ranks 135th on Noometry.