Z.ai (Zhipu), open weights

GLM-5.3

GLM-5.3 by Z.ai (Zhipu) ranks 26th of 354 ranked models on the Noometry Index as of October 2026, with a score of 54.8. Its strongest category is writing & preference, where it ranks 6th. API pricing starts at $1.40 per million input tokens and $4.40 per million output tokens, with a 1M-token context window.

Last verified

Specifications

Noometry rank
#26 of 354
Index score
54.8
Evidence
Confirmed 42 results
Released
August 14, 2026
Weights
Open weights
Reasoning
Yes
Context window
1M
Max output
131K
Input price
$1.40 / M
Output price
$4.40 / M
Blended price
$2.15 / M
Output speed
Not measured
Value
#141 of 219
Knowledge cutoff
Unknown
Input
text
Hugging Face
zai-org/GLM-5.3

Category scores

Each category score combines every public result we have in that category.

GLM-5.3 category scores
  1. Coding 59.5
  2. Agentic & Tool Use 36.4
  3. Reasoning 46.1
  4. Math 62.3
  5. Knowledge 58.3
  6. Multilingual 55.7
  7. Instruction Following 77.5
  8. Long Context 45.4
  9. Writing & Preference 75.7
GLM-5.3 category ranks
CategoryScoreRankResults
Coding59.5#148
Agentic & Tool Use36.4#381
Reasoning46.1#467
Math62.3#335
Knowledge58.3#373
Multilingual55.7#281
Instruction Following77.5#231
Long Context45.4#411
Writing & Preference75.7#64

Strengths and weaknesses

Categories where GLM-5.3 places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

GLM-5.3: strongest categories
CategoryScorevs medianRank
Writing & Preference75.7+21.9#6 of 312, top 2%
Coding59.5+20.8#14 of 340, top 5%
Instruction Following77.5+6.2#23 of 305, top 8%

Weakest categories

GLM-5.3: weakest categories
CategoryScorevs medianRank
Agentic & Tool Use36.4+6.0#38 of 154, top 25%
Long Context45.4+4.5#41 of 296, top 14%
Reasoning46.1+22.5#46 of 350, top 14%

Closest competitors

The models ranked just above and below GLM-5.3. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to GLM-5.3
ModelRankScoreBlended $/MSpeed
Qwen3.8 Max#2256.8$3—Compare
Gemini 3.1 Pro Preview#2356.7$4.50—Compare
Gemini 4 Argon#2456.5——Compare
Grok 4.5#2555.0$34Compare
Muse Spark 1.3#2754.8$2—Compare
Gemini 3 Pro#2854.8—1Compare
Claude Sonnet 5#2954.6$4—Compare
GPT-5.6 Luna#3054.6$0.4512Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

GLM-5.3 Coding benchmark results
BenchmarkScorePositionSettingSourceDate
DeepSWE69%#8 of 29, top 28%maxEpoch AI
FrontierCode40.1%#21 of 37, top 57%maxEpoch AI
CursorBench38%highEpoch AI
CursorBench33.3%lowEpoch AI
CursorBench42.6%#6 of 14, top 43%maxEpoch AI
LMArena WebDev1622#17 of 113, top 16%LMArena2026-10-08
FrontierSWE30.2%#9 of 18, top 50%maxEpoch AI
SciCode42%lowEpoch AI
SciCode59%#10 of 121, top 9%maxEpoch AI
WeirdML75.4%#15 of 119, top 13%maxEpoch AI
LMArena Coding1496#25 of 294, top 9%LMArena2026-10-08
ALE-Bench1,317#23 of 105, top 22%highEpoch AI

Agentic & Tool Use

GLM-5.3 Agentic & Tool Use benchmark results
BenchmarkScorePositionSettingSourceDate
APEX-Agents56.6%#16 of 49, top 33%Epoch AI
Vending-Bench 28,164#11 of 60, top 19%Epoch AI

Reasoning

GLM-5.3 Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
NYT Connections (extended)74.2%#46 of 91, top 51%high reasoningLech Mazur benchmarks
CritPt14.6%lowEpoch AI
CritPt19.1%#27 of 134, top 21%maxEpoch AI
Chess Puzzles21%#53 of 129, top 42%maxEpoch AI2026-08-24
LMArena Hard Prompts1489#16 of 297, top 6%LMArena2026-10-08
Mystery Game Puzzles33%#23 of 74, top 32%maxEpoch AI2026-08-30
DTBench87.7%#49 of 151, top 33%Epoch AI
LMCA55.5%#10 of 125, top 8%Epoch AI
Bench to the Future 30.15Best of 10Epoch AI
Epoch Capabilities Index155.61#27 of 213, top 13%Epoch AI2026-08-14

Math

GLM-5.3 Math benchmark results
BenchmarkScorePositionSettingSourceDate
FrontierMath (Tiers 1-3)68.8%#25 of 81, top 31%maxEpoch AI2026-08-25
FrontierMath Tier 429.3%#31 of 63, top 50%maxEpoch AI2026-08-25
OTIS Mock AIME 2024-202591.1%#49 of 173, top 29%maxEpoch AI2026-08-24
ProofBench49%#31 of 77, top 41%maxEpoch AI
LMArena Math1489#19 of 285, top 7%LMArena2026-10-08

Knowledge

GLM-5.3 Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
GPQA Diamond90.9%#29 of 186, top 16%maxEpoch AI2026-08-24
SimpleQA Verified41%#41 of 77, top 54%maxEpoch AI2026-08-28
LMArena Expert1516#14 of 273, top 6%LMArena2026-10-08

Multilingual

GLM-5.3 Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1457#28 of 297, top 10%LMArena2026-10-08
LMArena Chinese1528#19 of 285, top 7%LMArena2026-10-08
LMArena French1499#12 of 223, top 6%LMArena2026-10-08
LMArena German1499#6 of 231, top 3%LMArena2026-10-08
LMArena Japanese1453#22 of 211, top 11%LMArena2026-10-08
LMArena Korean1472#6 of 213, top 3%LMArena2026-10-08
LMArena Russian1463#31 of 283, top 11%LMArena2026-10-08
LMArena Spanish1460#35 of 226, top 16%LMArena2026-10-08

Instruction Following

GLM-5.3 Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Instruction Following1477#20 of 298, top 7%LMArena2026-10-08

Long Context

GLM-5.3 Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Longer Query1482#22 of 291, top 8%LMArena2026-10-08

Writing & Preference

GLM-5.3 Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1471#24 of 297, top 9%LMArena2026-10-08
LMArena Creative Writing1457#20 of 295, top 7%LMArena2026-10-08
EQ-Bench Creative Writing2075#6 of 115, top 6%EQ-Bench
LMArena Multi-Turn1472#28 of 295, top 10%LMArena2026-10-08

API pricing by provider

GLM-5.3 API prices
RouteInput $/MOutput $/MCached input $/MChecked
bedrock$1.68$5.28$0.312026-10-10
deepinfra$0.90$4$0.202026-10-10
fireworks$1.40$4.40$0.262026-10-10
mistral$1.40$4.40$0.142026-10-10
openrouter$0.039$4.80$0.0382026-10-10
together$1.40$4.40$0.262026-10-10
zai$1.40$4.40$0.262026-10-10

Compare GLM-5.3

Other Z.ai (Zhipu) models

Frequently asked questions

How good is GLM-5.3?

GLM-5.3 by Z.ai (Zhipu) ranks 26th of 354 ranked models on the Noometry Index as of October 2026, with a score of 54.8. Its strongest category is writing & preference, where it ranks 6th. API pricing starts at $1.40 per million input tokens and $4.40 per million output tokens, with a 1M-token context window.

How much does GLM-5.3 cost?

GLM-5.3 costs $1.40 per million input tokens and $4.40 per million output tokens on Z.ai (Zhipu)'s own API, with cached input at $0.26.

What is GLM-5.3's context window?

GLM-5.3 accepts up to 1M tokens of input and can write up to 131K tokens in one response.

Is GLM-5.3 open source?

Yes. GLM-5.3's weights are downloadable from Hugging Face (zai-org/GLM-5.3); check the license for commercial terms.

What are GLM-5.3's strengths and weaknesses?

Relative to other ranked models, GLM-5.3 places best in writing & preference, coding, instruction following and lowest in agentic & tool use, long context, reasoning.

What is GLM-5.3 best at?

Its best category is writing & preference, where it ranks 6th on Noometry.