Model comparison

GLM-5V-Turbo vs Hy4 preview

Hy4 preview is the stronger model overall, scoring 45.3 to 43.8 on the Noometry Index.

Last verified . 1 shared benchmarks.

GLM-5V-Turbo Z.ai (Zhipu)

43.8

Rank #84 Confirmed

Hy4 preview Tencent

45.3

Rank #73 Reported

Summary

  • They share 1 benchmark with published results for both. GLM-5V-Turbo scores higher in 0 categories and Hy4 preview in 3 categories; 3 gaps are clear of the uncertainty.
  • The widest gap is in math, where Hy4 preview leads 55.7 to 39.4.
  • Hy4 preview is cheaper at $0.75 / $2.25 per million input/output tokens, against $1.20 / $4 for GLM-5V-Turbo.
  • Hy4 preview accepts more context: 1.05M tokens versus 200K.
  • Hy4 preview has downloadable open weights; the other is API-only.

Side by side

GLM-5V-Turbo and Hy4 preview specifications
GLM-5V-TurboHy4 preview
ProviderZ.ai (Zhipu)Tencent
Noometry Index43.845.3
Released2026-04-012026-08-28
WeightsProprietaryOpen
Context window200K1.05M
Max output131K64K
Input $ / M tokens$1.20$0.75
Output $ / M tokens$4$2.25
Results tracked193

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hy4 preview leads

GLM-5V-Turbo: 42.1 (#111), Hy4 preview: 51.6 (#38)

Coding benchmarks
BenchmarkGLM-5V-TurboHy4 preview
LMArena WebDev14011632
LMArena Coding1466—

Reasoning Hy4 preview leads

GLM-5V-Turbo: 29.7 (#89), Hy4 preview: 31.9 (#79)

Reasoning benchmarks
BenchmarkGLM-5V-TurboHy4 preview
NYT Connections (extended)—68.2%
LMArena Hard Prompts1443—

Math Hy4 preview leads

GLM-5V-Turbo: 39.4 (#106), Hy4 preview: 55.7 (#42)

Math benchmarks
BenchmarkGLM-5V-TurboHy4 preview
ProofBench—75%
LMArena Math1441—

Knowledge Not comparable

GLM-5V-Turbo: 40.6 (#117), Hy4 preview: —

Knowledge benchmarks
BenchmarkGLM-5V-TurboHy4 preview
LMArena Expert1452—

Multimodal Not comparable

GLM-5V-Turbo: 40.9 (#42), Hy4 preview: —

Multimodal benchmarks
BenchmarkGLM-5V-TurboHy4 preview
LMArena Vision1264—
LMArena Document1416—

Multilingual Not comparable

GLM-5V-Turbo: 53.0 (#73), Hy4 preview: —

Multilingual benchmarks
BenchmarkGLM-5V-TurboHy4 preview
LMArena Non-English1420—
LMArena Chinese1488—
LMArena French1444—
LMArena German1423—
LMArena Korean1396—
LMArena Russian1431—
LMArena Spanish1450—

Instruction Following Not comparable

GLM-5V-Turbo: 75.0 (#80), Hy4 preview: —

Instruction Following benchmarks
BenchmarkGLM-5V-TurboHy4 preview
LMArena Instruction Following1423—

Long Context Not comparable

GLM-5V-Turbo: 44.0 (#80), Hy4 preview: —

Long Context benchmarks
BenchmarkGLM-5V-TurboHy4 preview
LMArena Longer Query1438—

Writing & Preference Not comparable

GLM-5V-Turbo: 62.5 (#73), Hy4 preview: —

Writing & Preference benchmarks
BenchmarkGLM-5V-TurboHy4 preview
LMArena Text1437—
LMArena Creative Writing1416—
LMArena Multi-Turn1432—

Frequently asked questions

Is GLM-5V-Turbo better than Hy4 preview?

Hy4 preview is the stronger model overall, scoring 45.3 to 43.8 on the Noometry Index.

Which is cheaper, GLM-5V-Turbo or Hy4 preview?

Hy4 preview is cheaper. It lists at $0.75 per million input tokens and $2.25 per million output tokens; GLM-5V-Turbo lists at $1.20 and $4.

Is GLM-5V-Turbo or Hy4 preview better for coding?

Hy4 preview scores higher on coding benchmarks: 51.6 versus 42.1 in the Noometry coding category.

Which has the bigger context window?

Hy4 preview does, with 1.05M tokens against 200K.

How many benchmarks do GLM-5V-Turbo and Hy4 preview share?

1 benchmark has published results for both models. GLM-5V-Turbo has 19 scored results on Noometry and Hy4 preview has 3.

Related comparisons

Go deeper