Model comparison

GLM-4.5V vs Hy3

Hy3 is the stronger model overall, scoring 44.2 to 39.8 on the Noometry Index.

Last verified . 13 shared benchmarks.

GLM-4.5V Z.ai (Zhipu)

39.8

Rank #158 Confirmed

Hy3 Tencent

44.2

Rank #79 Confirmed

Summary

  • They share 13 benchmarks with published results for both. GLM-4.5V scores higher in 1 category and Hy3 in 7 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Hy3 leads 62.2 to 52.5.
  • Hy3 is cheaper at $0.13 / $0.53 per million input/output tokens, against $0.60 / $1.80 for GLM-4.5V.
  • Hy3 accepts more context: 262K tokens versus 64K.

Side by side

GLM-4.5V and Hy3 specifications
GLM-4.5VHy3
ProviderZ.ai (Zhipu)Tencent
Noometry Index39.844.2
Released2025-08-112026-07-06
WeightsOpenOpen
Context window64K262K
Max output16K128K
Input $ / M tokens$0.60$0.13
Output $ / M tokens$1.80$0.53
Results tracked1519

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hy3 leads

GLM-4.5V: 39.5 (#155), Hy3: 46.8 (#63)

Coding benchmarks
BenchmarkGLM-4.5VHy3
LMArena Coding13471464
LMArena WebDev—1508

Reasoning GLM-4.5V leads

GLM-4.5V: 27.4 (#119), Hy3: 26.1 (#136)

Reasoning benchmarks
BenchmarkGLM-4.5VHy3
LMArena Hard Prompts13341447
Kagi LLM Benchmark59.8%—
NYT Connections (extended)—41.2%

Math Hy3 leads

GLM-4.5V: 37.4 (#159), Hy3: 40.1 (#93)

Math benchmarks
BenchmarkGLM-4.5VHy3
LMArena Math13541475

Knowledge Hy3 leads

GLM-4.5V: 37.5 (#156), Hy3: 40.8 (#114)

Knowledge benchmarks
BenchmarkGLM-4.5VHy3
LMArena Expert13531460

Multimodal Not comparable

GLM-4.5V: 34.3 (#92), Hy3: —

Multimodal benchmarks
BenchmarkGLM-4.5VHy3
LMArena Vision1154—

Multilingual Hy3 leads

GLM-4.5V: 44.6 (#177), Hy3: 53.5 (#65)

Multilingual benchmarks
BenchmarkGLM-4.5VHy3
LMArena Non-English13031426
LMArena Chinese13371493
LMArena Russian12981432
LMArena Spanish13361456
LMArena French—1461
LMArena German—1439
LMArena Japanese—1392
LMArena Korean—1395

Instruction Following Hy3 leads

GLM-4.5V: 69.2 (#175), Hy3: 75.1 (#70)

Instruction Following benchmarks
BenchmarkGLM-4.5VHy3
LMArena Instruction Following13111426

Long Context Hy3 leads

GLM-4.5V: 39.6 (#171), Hy3: 44.1 (#75)

Long Context benchmarks
BenchmarkGLM-4.5VHy3
LMArena Longer Query13041442

Writing & Preference Hy3 leads

GLM-4.5V: 52.5 (#170), Hy3: 62.2 (#81)

Writing & Preference benchmarks
BenchmarkGLM-4.5VHy3
LMArena Text13331439
LMArena Creative Writing12951402
LMArena Multi-Turn13321436

Frequently asked questions

Is GLM-4.5V better than Hy3?

Hy3 is the stronger model overall, scoring 44.2 to 39.8 on the Noometry Index.

Which is cheaper, GLM-4.5V or Hy3?

Hy3 is cheaper. It lists at $0.13 per million input tokens and $0.53 per million output tokens; GLM-4.5V lists at $0.60 and $1.80.

Is GLM-4.5V or Hy3 better for coding?

Hy3 scores higher on coding benchmarks: 46.8 versus 39.5 in the Noometry coding category.

Which has the bigger context window?

Hy3 does, with 262K tokens against 64K.

How many benchmarks do GLM-4.5V and Hy3 share?

13 benchmarks have published results for both models. GLM-4.5V has 15 scored results on Noometry and Hy3 has 19.

Related comparisons

Go deeper