Z.ai (Zhipu), open weights

GLM-4.5

GLM-4.5 by Z.ai (Zhipu) ranks 122nd of 354 ranked models on the Noometry Index as of October 2026, with a score of 42.0. Its strongest category is multilingual, where it ranks 77th. API pricing starts at $0.60 per million input tokens and $2.20 per million output tokens, with a 131K-token context window.

Last verified

Specifications

Noometry rank
#122 of 354
Index score
42.0
Evidence
Confirmed 27 results
Released
July 27, 2025
Weights
Open weights
Reasoning
Yes
Context window
131K
Max output
98K
Input price
$0.60 / M
Output price
$2.20 / M
Blended price
$1 / M
Output speed
32 tokens/s Kagi
Value
#109 of 219
Knowledge cutoff
April 2025
Input
text
Hugging Face
zai-org/GLM-4.5

Category scores

Each category score combines every public result we have in that category.

GLM-4.5 category scores
  1. Coding 41.4
  2. Reasoning 28.6
  3. Math 39.0
  4. Knowledge 35.9
  5. Multilingual 52.8
  6. Instruction Following 74.1
  7. Long Context 38.2
  8. Writing & Preference 57.5
GLM-4.5 category ranks
CategoryScoreRankResults
Coding41.4#1253
Reasoning28.6#1002
Math39.0#1161
Knowledge35.9#1793
Multilingual52.8#771
Instruction Following74.1#1041
Long Context38.2#2012
Writing & Preference57.5#1275

Strengths and weaknesses

Categories where GLM-4.5 places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

GLM-4.5: strongest categories
CategoryScorevs medianRank
Multilingual52.8+5.4#77 of 297, top 26%
Reasoning28.6+5.0#100 of 350, top 29%
Instruction Following74.1+2.8#104 of 305, top 35%

Weakest categories

GLM-4.5: weakest categories
CategoryScorevs medianRank
Long Context38.2−2.7#201 of 296, top 68%
Knowledge35.9−1.5#179 of 314, top 58%
Writing & Preference57.5+3.7#127 of 312, top 41%

Closest competitors

The models ranked just above and below GLM-4.5. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to GLM-4.5
ModelRankScoreBlended $/MSpeed
Amazon Nova Experimental Chat 10 20#11842.1——Compare
Qwen3.5 122B-A10B#11942.1$1.10—Compare
Longcat Flash Chat#12042.1—69Compare
Solar Pro4#12142.1$0.52—Compare
Qwen3.5 35B-A3B#12342.0$0.69—Compare
GLM-4.7#12442.0$1—Compare
GPT-5.4 nano#12541.9$0.4619Compare
Amazon Nova Experimental Chat 10 09#12641.9——Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

GLM-4.5 Coding benchmark results
BenchmarkScorePositionSettingSourceDate
SWE-bench Verified (bash only)54.2%#25 of 39, top 65%SWE-bench2025-08-22
WeirdML40.6%#76 of 119, top 64%thinkingEpoch AI
LMArena Coding1434#105 of 294, top 36%LMArena2026-10-08
ALE-Bench344.82#94 of 105, top 90%Epoch AI
AlgoTune1.52#10 of 18, top 56%thinkingEpoch AI

Reasoning

GLM-4.5 Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
Kagi LLM Benchmark57.9%#45 of 99, top 46%Kagi LLM Benchmark
LMArena Hard Prompts1429#90 of 297, top 31%LMArena2026-10-08

Math

GLM-4.5 Math benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Math1427#91 of 285, top 32%LMArena2026-10-08

Knowledge

GLM-4.5 Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
Humanity's Last Exam8.3%#26 of 41, top 64%Epoch AI
Confabulations (lower is better)11.3%#4 of 51, top 8%Lech Mazur benchmarks
LMArena Expert1433#92 of 273, top 34%LMArena2026-10-08

Multilingual

GLM-4.5 Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1417#77 of 297, top 26%LMArena2026-10-08
LMArena Chinese1465#79 of 285, top 28%LMArena2026-10-08
LMArena French1418#104 of 223, top 47%LMArena2026-10-08
LMArena German1407#89 of 231, top 39%LMArena2026-10-08
LMArena Japanese1415#45 of 211, top 22%LMArena2026-10-08
LMArena Korean1380#78 of 213, top 37%LMArena2026-10-08
LMArena Russian1414#90 of 283, top 32%LMArena2026-10-08
LMArena Spanish1454#44 of 226, top 20%LMArena2026-10-08

Instruction Following

GLM-4.5 Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Instruction Following1404#98 of 298, top 33%LMArena2026-10-08

Long Context

GLM-4.5 Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
Fiction.LiveBench58.3%#30 of 47, top 64%Epoch AI
LMArena Longer Query1412#104 of 291, top 36%LMArena2026-10-08

Writing & Preference

GLM-4.5 Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1430#78 of 297, top 27%LMArena2026-10-08
LMArena Creative Writing1395#87 of 295, top 30%LMArena2026-10-08
Short-Story Creative Writing73.4%#29 of 39, top 75%Epoch AI
EQ-Bench Creative Writing1343#75 of 115, top 66%EQ-Bench
LMArena Multi-Turn1415#102 of 295, top 35%LMArena2026-10-08

API pricing by provider

GLM-4.5 API prices
RouteInput $/MOutput $/MCached input $/MChecked
openrouter$0.60$2.20$0.112026-10-10
zai$0.60$2.20$0.112026-10-10

Compare GLM-4.5

Other Z.ai (Zhipu) models

Frequently asked questions

How good is GLM-4.5?

GLM-4.5 by Z.ai (Zhipu) ranks 122nd of 354 ranked models on the Noometry Index as of October 2026, with a score of 42.0. Its strongest category is multilingual, where it ranks 77th. API pricing starts at $0.60 per million input tokens and $2.20 per million output tokens, with a 131K-token context window.

How much does GLM-4.5 cost?

GLM-4.5 costs $0.60 per million input tokens and $2.20 per million output tokens on Z.ai (Zhipu)'s own API, with cached input at $0.11.

What is GLM-4.5's context window?

GLM-4.5 accepts up to 131K tokens of input and can write up to 98K tokens in one response.

Is GLM-4.5 open source?

Yes. GLM-4.5's weights are downloadable from Hugging Face (zai-org/GLM-4.5); check the license for commercial terms.

How fast is GLM-4.5?

GLM-4.5 generated about 32 output tokens per second in the Kagi LLM Benchmark's timed runs. Speed varies by provider, load and reasoning effort.

What are GLM-4.5's strengths and weaknesses?

Relative to other ranked models, GLM-4.5 places best in multilingual, reasoning, instruction following and lowest in long context, knowledge, writing & preference.

What is GLM-4.5 best at?

Its best category is multilingual, where it ranks 77th on Noometry.