Model comparison
GLM-5V-Turbo vs Hy3
GLM-5V-Turbo and Hy3 score almost the same on the Noometry Index (43.8 vs 44.2), so choose on price, context window or the category you care about most.
Last verified . 17 shared benchmarks.
Summary
- They share 17 benchmarks with published results for both. GLM-5V-Turbo scores higher in 2 categories and Hy3 in 6 categories; 2 gaps are clear of the uncertainty.
- The widest gap is in coding, where Hy3 leads 46.8 to 42.1.
- Hy3 is cheaper at $0.0825 / $0.33 per million input/output tokens, against $1.20 / $4 for GLM-5V-Turbo.
- Hy3 accepts more context: 262K tokens versus 200K.
- Hy3 has downloadable open weights; the other is API-only.
Side by side
| GLM-5V-Turbo | Hy3 | |
|---|---|---|
| Provider | Z.ai (Zhipu) | Tencent |
| Noometry Index | 43.8 | 44.2 |
| Released | 2026-04-01 | 2026-07-06 |
| Weights | Proprietary | Open |
| Context window | 200K | 262K |
| Max output | 131K | 128K |
| Input $ / M tokens | $1.20 | $0.0825 |
| Output $ / M tokens | $4 | $0.33 |
| Results tracked | 19 | 19 |
Sponsored placements are available on pages like this one. Advertise on Noometry
Category by category
Coding Hy3 leads
GLM-5V-Turbo: 42.1 (#111), Hy3: 46.8 (#63)
| Benchmark | GLM-5V-Turbo | Hy3 |
|---|---|---|
| LMArena WebDev | 1401 | 1508 |
| LMArena Coding | 1466 | 1464 |
Reasoning GLM-5V-Turbo leads
GLM-5V-Turbo: 29.7 (#89), Hy3: 26.1 (#136)
| Benchmark | GLM-5V-Turbo | Hy3 |
|---|---|---|
| LMArena Hard Prompts | 1443 | 1447 |
| NYT Connections (extended) | — | 41.2% |
Math Too close to call
GLM-5V-Turbo: 39.4 (#106), Hy3: 40.1 (#93)
| Benchmark | GLM-5V-Turbo | Hy3 |
|---|---|---|
| LMArena Math | 1441 | 1475 |
Knowledge Too close to call
GLM-5V-Turbo: 40.6 (#117), Hy3: 40.8 (#114)
| Benchmark | GLM-5V-Turbo | Hy3 |
|---|---|---|
| LMArena Expert | 1452 | 1460 |
Multimodal Not comparable
GLM-5V-Turbo: 40.9 (#42), Hy3: —
| Benchmark | GLM-5V-Turbo | Hy3 |
|---|---|---|
| LMArena Vision | 1264 | — |
| LMArena Document | 1416 | — |
Multilingual Too close to call
GLM-5V-Turbo: 53.0 (#73), Hy3: 53.5 (#65)
| Benchmark | GLM-5V-Turbo | Hy3 |
|---|---|---|
| LMArena Non-English | 1420 | 1426 |
| LMArena Chinese | 1488 | 1493 |
| LMArena French | 1444 | 1461 |
| LMArena German | 1423 | 1439 |
| LMArena Korean | 1396 | 1395 |
| LMArena Russian | 1431 | 1432 |
| LMArena Spanish | 1450 | 1456 |
| LMArena Japanese | — | 1392 |
Instruction Following Too close to call
GLM-5V-Turbo: 75.0 (#80), Hy3: 75.1 (#70)
| Benchmark | GLM-5V-Turbo | Hy3 |
|---|---|---|
| LMArena Instruction Following | 1423 | 1426 |
Long Context Too close to call
GLM-5V-Turbo: 44.0 (#80), Hy3: 44.1 (#75)
| Benchmark | GLM-5V-Turbo | Hy3 |
|---|---|---|
| LMArena Longer Query | 1438 | 1442 |
Writing & Preference Too close to call
GLM-5V-Turbo: 62.5 (#73), Hy3: 62.2 (#81)
| Benchmark | GLM-5V-Turbo | Hy3 |
|---|---|---|
| LMArena Text | 1437 | 1439 |
| LMArena Creative Writing | 1416 | 1402 |
| LMArena Multi-Turn | 1432 | 1436 |
Frequently asked questions
Is GLM-5V-Turbo better than Hy3?
GLM-5V-Turbo and Hy3 score almost the same on the Noometry Index (43.8 vs 44.2), so choose on price, context window or the category you care about most.
Which is cheaper, GLM-5V-Turbo or Hy3?
Hy3 is cheaper. It lists at $0.0825 per million input tokens and $0.33 per million output tokens; GLM-5V-Turbo lists at $1.20 and $4.
Is GLM-5V-Turbo or Hy3 better for coding?
Hy3 scores higher on coding benchmarks: 46.8 versus 42.1 in the Noometry coding category.
Which has the bigger context window?
Hy3 does, with 262K tokens against 200K.
How many benchmarks do GLM-5V-Turbo and Hy3 share?
17 benchmarks have published results for both models. GLM-5V-Turbo has 19 scored results on Noometry and Hy3 has 19.