Z.ai (Zhipu), open weights
GLM-4.6
GLM-4.6 by Z.ai (Zhipu) ranks 135th of 354 ranked models on the Noometry Index as of October 2026, with a score of 41.4. Its strongest category is agentic & tool use, where it ranks 66th. API pricing starts at $0.60 per million input tokens and $2.20 per million output tokens, with a 205K-token context window.
Last verified
Specifications
- Noometry rank
- #135 of 354
- Index score
- 41.4
- Evidence
- Confirmed 29 results
- Provider
- Z.ai (Zhipu)
- Released
- September 30, 2025
- Weights
- Open weights
- Reasoning
- Yes
- Context window
- 205K
- Max output
- 131K
- Input price
- $0.60 / M
- Output price
- $2.20 / M
- Blended price
- $1 / M
- Output speed
- 12 tokens/s Kagi
- Value
- #113 of 219
- Knowledge cutoff
- April 2025
- Input
- text
- Hugging Face
- zai-org/GLM-4.6
Category scores
Each category score combines every public result we have in that category.
- Coding 40.1
- Agentic & Tool Use 32.3
- Reasoning 23.7
- Math 39.1
- Knowledge 40.2
- Multilingual 53.5
- Instruction Following 74.3
- Long Context 43.4
- Writing & Preference 61.1
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 40.1 | #148 | 4 |
| Agentic & Tool Use | 32.3 | #66 | 2 |
| Reasoning | 23.7 | #172 | 3 |
| Math | 39.1 | #111 | 1 |
| Knowledge | 40.2 | #124 | 2 |
| Multilingual | 53.5 | #66 | 1 |
| Instruction Following | 74.3 | #98 | 1 |
| Long Context | 43.4 | #94 | 1 |
| Writing & Preference | 61.1 | #90 | 4 |
Strengths and weaknesses
Categories where GLM-4.6 places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Multilingual | 53.5 | +6.1 | #66 of 297, top 23% |
| Writing & Preference | 61.1 | +7.3 | #90 of 312, top 29% |
| Long Context | 43.4 | +2.5 | #94 of 296, top 32% |
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Reasoning | 23.7 | +0.1 | #172 of 350, top 50% |
| Coding | 40.1 | +1.4 | #148 of 340, top 44% |
| Agentic & Tool Use | 32.3 | +2.0 | #66 of 154, top 43% |
Closest competitors
The models ranked just above and below GLM-4.6. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Muse Glimmer | #131 | 41.7 | — | — | Compare |
| o4-mini | #132 | 41.6 | $1.93 | 6 | Compare |
| Gemini 3.5 Flash Lite | #133 | 41.5 | $0.85 | — | Compare |
| Grok 4.1 | #134 | 41.5 | — | — | Compare |
| Grok 4.1 Fast | #136 | 41.4 | $0.28 | — | Compare |
| GLM-4.6V | #137 | 41.3 | $0.45 | — | Compare |
| MiMo-V2-Flash | #138 | 41.3 | $0.18 | — | Compare |
| Hunyuan Turbos 20250226 | #139 | 41.3 | — | — | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| SWE-bench Verified (bash only) | 55.4% | #23 of 39, top 59% | SWE-bench | 2025-12-01 | |
| LMArena WebDev | 1340 | #90 of 113, top 80% | LMArena | 2026-10-08 | |
| SciCode | 38.4% | #88 of 121, top 73% | Epoch AI | ||
| LMArena Coding | 1449 | #87 of 294, top 30% | LMArena | 2026-10-08 | |
| ALE-Bench | 340.82 | #95 of 105, top 91% | Epoch AI |
Agentic & Tool Use
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| Terminal-Bench | 24.5% | #35 of 41, top 86% | Epoch AI | ||
| Berkeley Function Calling Leaderboard | 72.4% | #4 of 49, top 9% | fc thinking | Berkeley Function Calling Leaderboard |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| Kagi LLM Benchmark | 45.7% | Kagi LLM Benchmark | |||
| Kagi LLM Benchmark | 47.4% | #69 of 99, top 70% | Kagi LLM Benchmark | ||
| CritPt | 1.1% | #80 of 134, top 60% | Epoch AI | ||
| LMArena Hard Prompts | 1440 | #81 of 297, top 28% | LMArena | 2026-10-08 |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Math | 1432 | #80 of 285, top 29% | LMArena | 2026-10-08 | |
| FrontierMath (Feb 2025 set) | 3.8% | #51 of 68, top 75% | Epoch AI | 2025-12-08 | |
| FrontierMath Tier 4 (v1) | 2.1% | #36 of 55, top 66% | Epoch AI | 2025-12-08 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| Vectara Hallucination Rate (lower is better) | 9.5% | #50 of 96, top 53% | Vectara Hallucination Leaderboard | ||
| LMArena Expert | 1431 | #95 of 273, top 35% | LMArena | 2026-10-08 |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1426 | #66 of 297, top 23% | LMArena | 2026-10-08 | |
| LMArena Chinese | 1499 | #45 of 285, top 16% | LMArena | 2026-10-08 | |
| LMArena French | 1459 | #53 of 223, top 24% | LMArena | 2026-10-08 | |
| LMArena German | 1447 | #48 of 231, top 21% | LMArena | 2026-10-08 | |
| LMArena Japanese | 1393 | #66 of 211, top 32% | LMArena | 2026-10-08 | |
| LMArena Korean | 1400 | #53 of 213, top 25% | LMArena | 2026-10-08 | |
| LMArena Russian | 1419 | #83 of 283, top 30% | LMArena | 2026-10-08 | |
| LMArena Spanish | 1436 | #71 of 226, top 32% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Instruction Following | 1410 | #87 of 298, top 30% | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 1422 | #87 of 291, top 30% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1440 | #62 of 297, top 21% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1411 | #62 of 295, top 22% | LMArena | 2026-10-08 | |
| EQ-Bench Creative Writing | 1411 | #71 of 115, top 62% | EQ-Bench | ||
| LMArena Multi-Turn | 1427 | #87 of 295, top 30% | LMArena | 2026-10-08 |
API pricing by provider
| Route | Input $/M | Output $/M | Cached input $/M | Checked |
|---|---|---|---|---|
| deepinfra | $0.50 | $2 | $0.10 | 2026-10-10 |
| openrouter | $0.50 | $2 | $0.10 | 2026-10-10 |
| zai | $0.60 | $2.20 | $0.11 | 2026-10-10 |
Compare GLM-4.6
- GLM-4.6 vs GLM-4.5V
- GLM-4.6 vs Grok 4.1
- GLM-4.6 vs Grok 4.1 Fast
- GLM-4.6 vs Gemini 3.5 Flash Lite
- GLM-4.6 vs GLM-4.6V
- GLM-4.6 vs o4-mini
- GLM-4.6 vs MiMo-V2-Flash
- GLM-4.6 vs GPT-6 Astra
- GLM-4.6 vs Claude Fable 5.1
- GLM-4.6 vs Gemini 3.8 Flash
- GLM-4.6 vs Kimi K3
- GLM-4.6 vs Grok 4.6
- GLM-4.6 vs Qwen3.8 Max
- GLM-4.6 vs Muse Spark 1.3
Other Z.ai (Zhipu) models
- GLM-5.354.8
- GLM-5.3-Flash51.8
- GLM-5.251.1
- GLM-5.147.8
- GLM-546.1
- GLM-5V-Turbo43.8
- GLM-4.542.0
- GLM-4.742.0
Frequently asked questions
How good is GLM-4.6?
GLM-4.6 by Z.ai (Zhipu) ranks 135th of 354 ranked models on the Noometry Index as of October 2026, with a score of 41.4. Its strongest category is agentic & tool use, where it ranks 66th. API pricing starts at $0.60 per million input tokens and $2.20 per million output tokens, with a 205K-token context window.
How much does GLM-4.6 cost?
GLM-4.6 costs $0.60 per million input tokens and $2.20 per million output tokens on Z.ai (Zhipu)'s own API, with cached input at $0.11.
What is GLM-4.6's context window?
GLM-4.6 accepts up to 205K tokens of input and can write up to 131K tokens in one response.
Is GLM-4.6 open source?
Yes. GLM-4.6's weights are downloadable from Hugging Face (zai-org/GLM-4.6); check the license for commercial terms.
How fast is GLM-4.6?
GLM-4.6 generated about 12 output tokens per second in the Kagi LLM Benchmark's timed runs. Speed varies by provider, load and reasoning effort.
What are GLM-4.6's strengths and weaknesses?
Relative to other ranked models, GLM-4.6 places best in multilingual, writing & preference, long context and lowest in reasoning, coding, agentic & tool use.
What is GLM-4.6 best at?
Its best category is agentic & tool use, where it ranks 66th on Noometry.