Z.ai (Zhipu), open weights

GLM-5.3-Flash

GLM-5.3-Flash by Z.ai (Zhipu) ranks 41st of 354 ranked models on the Noometry Index as of October 2026, with a score of 51.8. Its strongest category is instruction following, where it ranks 20th. API pricing starts at $0.15 per million input tokens and $0.50 per million output tokens, with a 1M-token context window.

Last verified

Specifications

Noometry rank
#41 of 354
Index score
51.8
Evidence
Confirmed 40 results
Released
August 20, 2026
Weights
Open weights
Reasoning
Yes
Context window
1M
Max output
131K
Input price
$0.15 / M
Output price
$0.50 / M
Blended price
$0.24 / M
Output speed
Not measured
Value
#39 of 219
Knowledge cutoff
Unknown
Input
text, image, video, pdf

Category scores

Each category score combines every public result we have in that category.

GLM-5.3-Flash category scores
  1. Coding 53.1
  2. Agentic & Tool Use 34.2
  3. Reasoning 48.0
  4. Math 53.3
  5. Knowledge 58.4
  6. Multimodal 42.8
  7. Multilingual 56.0
  8. Instruction Following 77.5
  9. Long Context 45.4
  10. Writing & Preference 65.3
GLM-5.3-Flash category ranks
CategoryScoreRankResults
Coding53.1#317
Agentic & Tool Use34.2#472
Reasoning48.0#427
Math53.3#475
Knowledge58.4#362
Multimodal42.8#271
Multilingual56.0#251
Instruction Following77.5#201
Long Context45.4#391
Writing & Preference65.3#503

Strengths and weaknesses

Categories where GLM-5.3-Flash places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

GLM-5.3-Flash: strongest categories
CategoryScorevs medianRank
Instruction Following77.5+6.3#20 of 305, top 7%
Multilingual56.0+8.6#25 of 297, top 9%
Coding53.1+14.4#31 of 340, top 10%

Weakest categories

GLM-5.3-Flash: weakest categories
CategoryScorevs medianRank
Agentic & Tool Use34.2+3.9#47 of 154, top 31%
Multimodal42.8+4.3#27 of 128, top 22%
Writing & Preference65.3+11.5#50 of 312, top 17%

Closest competitors

The models ranked just above and below GLM-5.3-Flash. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to GLM-5.3-Flash
ModelRankScoreBlended $/MSpeed
Grok 4.7#3753.1$3—Compare
DeepSeek V4.1 Flash#3852.8$0.26—Compare
GPT-5.2 Pro#3952.3$57.75—Compare
Gemini 3 Flash Preview#4052.3$1.13—Compare
Qwen3.7 Max#4251.5$3.75—Compare
Qwen3.6 Max Preview#4351.5$2.92—Compare
GLM-5.2#4451.1$2.1523Compare
GPT-5#4550.9$3.442Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

GLM-5.3-Flash Coding benchmark results
BenchmarkScorePositionSettingSourceDate
DeepSWE63.4%#16 of 29, top 56%maxEpoch AI
FrontierCode31.8%#25 of 37, top 68%maxEpoch AI
CursorBench31.1%highEpoch AI
CursorBench26.9%lowEpoch AI
CursorBench36.8%#12 of 14, top 86%maxEpoch AI
LMArena WebDev1609#21 of 113, top 19%LMArena2026-10-08
FrontierSWE18.1%#15 of 18, top 84%maxEpoch AI
SciCode51.6%#37 of 121, top 31%Epoch AI
LMArena Coding1508#13 of 294, top 5%LMArena2026-10-08
ALE-Bench303.55#98 of 105, top 94%highEpoch AI

Agentic & Tool Use

GLM-5.3-Flash Agentic & Tool Use benchmark results
BenchmarkScorePositionSettingSourceDate
APEX-Agents52.8%#22 of 49, top 45%Epoch AI
GDP.pdf14%#30 of 36, top 84%maxEpoch AI

Reasoning

GLM-5.3-Flash Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
ARC-AGI-250.1%highEpoch AI
ARC-AGI-227.9%lowEpoch AI
ARC-AGI-265.8%#23 of 83, top 28%maxEpoch AI
ARC-AGI-171.8%highEpoch AI
ARC-AGI-147%lowEpoch AI
ARC-AGI-191%#24 of 83, top 29%maxEpoch AI
CritPt15.4%#35 of 134, top 27%Epoch AI
Chess Puzzles14%#70 of 129, top 55%maxEpoch AI2026-08-26
LMArena Hard Prompts1491#15 of 297, top 6%LMArena2026-10-08
Mystery Game Puzzles8%#64 of 74, top 87%maxEpoch AI2026-08-28
Surface Evolver Bench52.5%#14 of 25, top 57%maxEpoch AI
Bench to the Future 30.15#2 of 10, top 20%Epoch AI
Epoch Capabilities Index151.88#46 of 213, top 22%Epoch AI2026-08-20

Math

GLM-5.3-Flash Math benchmark results
BenchmarkScorePositionSettingSourceDate
FrontierMath (Tiers 1-3)55.8%#41 of 81, top 51%maxEpoch AI2026-08-27
FrontierMath Tier 417.1%#45 of 63, top 72%maxEpoch AI2026-08-27
OTIS Mock AIME 2024-202593.9%#40 of 173, top 24%maxEpoch AI2026-08-26
ProofBench21%#47 of 77, top 62%maxEpoch AI
LMArena Math1500#12 of 285, top 5%LMArena2026-10-08

Knowledge

GLM-5.3-Flash Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
GPQA Diamond90.2%#38 of 186, top 21%maxEpoch AI2026-08-26
LMArena Expert1513#15 of 273, top 6%LMArena2026-10-08

Multimodal

GLM-5.3-Flash Multimodal benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Vision1296#19 of 122, top 16%LMArena2026-10-09

Multilingual

GLM-5.3-Flash Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1462#25 of 297, top 9%LMArena2026-10-08
LMArena Chinese1527#20 of 285, top 8%LMArena2026-10-08
LMArena French1496#15 of 223, top 7%LMArena2026-10-08
LMArena German1470#28 of 231, top 13%LMArena2026-10-08
LMArena Japanese1429#32 of 211, top 16%LMArena2026-10-08
LMArena Korean1446#21 of 213, top 10%LMArena2026-10-08
LMArena Russian1469#27 of 283, top 10%LMArena2026-10-08
LMArena Spanish1471#22 of 226, top 10%LMArena2026-10-08

Instruction Following

GLM-5.3-Flash Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Instruction Following1478#17 of 298, top 6%LMArena2026-10-08

Long Context

GLM-5.3-Flash Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Longer Query1482#21 of 291, top 8%LMArena2026-10-08

Writing & Preference

GLM-5.3-Flash Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1471#25 of 297, top 9%LMArena2026-10-08
LMArena Creative Writing1442#33 of 295, top 12%LMArena2026-10-08
LMArena Multi-Turn1467#33 of 295, top 12%LMArena2026-10-08

API pricing by provider

GLM-5.3-Flash API prices
RouteInput $/MOutput $/MCached input $/MChecked
deepinfra$0.15$0.50$0.032026-10-10
fireworks$0.15$0.50$0.032026-10-10
openrouter$0.15$0.50$0.032026-10-10
together$0.15$0.50$0.032026-10-10
zai$0.15$0.50$0.032026-10-10

Compare GLM-5.3-Flash

Other Z.ai (Zhipu) models

Frequently asked questions

How good is GLM-5.3-Flash?

GLM-5.3-Flash by Z.ai (Zhipu) ranks 41st of 354 ranked models on the Noometry Index as of October 2026, with a score of 51.8. Its strongest category is instruction following, where it ranks 20th. API pricing starts at $0.15 per million input tokens and $0.50 per million output tokens, with a 1M-token context window.

How much does GLM-5.3-Flash cost?

GLM-5.3-Flash costs $0.15 per million input tokens and $0.50 per million output tokens on Z.ai (Zhipu)'s own API, with cached input at $0.03.

What is GLM-5.3-Flash's context window?

GLM-5.3-Flash accepts up to 1M tokens of input and can write up to 131K tokens in one response.

Is GLM-5.3-Flash open source?

Yes. GLM-5.3-Flash's weights are downloadable from Hugging Face (zai-org/GLM-5.3-Flash); check the license for commercial terms.

What are GLM-5.3-Flash's strengths and weaknesses?

Relative to other ranked models, GLM-5.3-Flash places best in instruction following, multilingual, coding and lowest in agentic & tool use, multimodal, writing & preference.

What is GLM-5.3-Flash best at?

Its best category is instruction following, where it ranks 20th on Noometry.