Model comparison

GLM-4.5 vs MiMo-V2-Flash

GLM-4.5 and MiMo-V2-Flash score almost the same on the Noometry Index (42.0 vs 41.3), so choose on price, context window or the category you care about most.

Last verified . 18 shared benchmarks.

GLM-4.5 Z.ai (Zhipu)

42.0

Rank #122 Confirmed

MiMo-V2-Flash Xiaomi

41.3

Rank #138 Confirmed

Summary

  • They share 18 benchmarks with published results for both. GLM-4.5 scores higher in 5 categories and MiMo-V2-Flash in 3 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in coding, where GLM-4.5 leads 41.4 to 36.1.
  • MiMo-V2-Flash is cheaper at $0.14 / $0.28 per million input/output tokens, against $0.60 / $2.20 for GLM-4.5.
  • MiMo-V2-Flash accepts more context: 262K tokens versus 131K.

Side by side

GLM-4.5 and MiMo-V2-Flash specifications
GLM-4.5MiMo-V2-Flash
ProviderZ.ai (Zhipu)Xiaomi
Noometry Index42.041.3
Released2025-07-272025-12-16
WeightsOpenOpen
Context window131K262K
Max output98K66K
Input $ / M tokens$0.60$0.14
Output $ / M tokens$2.20$0.28
Results tracked2721

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding GLM-4.5 leads

GLM-4.5: 41.4 (#125), MiMo-V2-Flash: 36.1 (#211)

Coding benchmarks
BenchmarkGLM-4.5MiMo-V2-Flash
LMArena Coding14341443
ALE-Bench344.82737.95
SWE-bench Verified (bash only)54.2%—
LMArena WebDev—1330
SciCode—25.9%
WeirdML40.6%—
AlgoTune1.52—

Reasoning GLM-4.5 leads

GLM-4.5: 28.6 (#100), MiMo-V2-Flash: 24.9 (#157)

Reasoning benchmarks
BenchmarkGLM-4.5MiMo-V2-Flash
LMArena Hard Prompts14291420
Kagi LLM Benchmark57.9%—
CritPt—0%

Math Too close to call

GLM-4.5: 39.0 (#116), MiMo-V2-Flash: 38.3 (#139)

Math benchmarks
BenchmarkGLM-4.5MiMo-V2-Flash
LMArena Math14271396

Knowledge MiMo-V2-Flash leads

GLM-4.5: 35.9 (#179), MiMo-V2-Flash: 39.7 (#131)

Knowledge benchmarks
BenchmarkGLM-4.5MiMo-V2-Flash
LMArena Expert14331425
Humanity's Last Exam8.3%—
Confabulations11.3%—

Multilingual GLM-4.5 leads

GLM-4.5: 52.8 (#77), MiMo-V2-Flash: 51.0 (#113)

Multilingual benchmarks
BenchmarkGLM-4.5MiMo-V2-Flash
LMArena Non-English14171392
LMArena Chinese14651462
LMArena French14181429
LMArena German14071395
LMArena Japanese14151325
LMArena Korean13801358
LMArena Russian14141387
LMArena Spanish14541420

Instruction Following Too close to call

GLM-4.5: 74.1 (#104), MiMo-V2-Flash: 73.5 (#120)

Instruction Following benchmarks
BenchmarkGLM-4.5MiMo-V2-Flash
LMArena Instruction Following14041392

Long Context MiMo-V2-Flash leads

GLM-4.5: 38.2 (#201), MiMo-V2-Flash: 43.0 (#110)

Long Context benchmarks
BenchmarkGLM-4.5MiMo-V2-Flash
LMArena Longer Query14121409
Fiction.LiveBench58.3%—

Writing & Preference MiMo-V2-Flash leads

GLM-4.5: 57.5 (#127), MiMo-V2-Flash: 59.7 (#106)

Writing & Preference benchmarks
BenchmarkGLM-4.5MiMo-V2-Flash
LMArena Text14301411
LMArena Creative Writing13951375
LMArena Multi-Turn14151404
Short-Story Creative Writing73.4%—
EQ-Bench Creative Writing1343—

Frequently asked questions

Is GLM-4.5 better than MiMo-V2-Flash?

GLM-4.5 and MiMo-V2-Flash score almost the same on the Noometry Index (42.0 vs 41.3), so choose on price, context window or the category you care about most.

Which is cheaper, GLM-4.5 or MiMo-V2-Flash?

MiMo-V2-Flash is cheaper. It lists at $0.14 per million input tokens and $0.28 per million output tokens; GLM-4.5 lists at $0.60 and $2.20.

Is GLM-4.5 or MiMo-V2-Flash better for coding?

GLM-4.5 scores higher on coding benchmarks: 41.4 versus 36.1 in the Noometry coding category.

Which has the bigger context window?

MiMo-V2-Flash does, with 262K tokens against 131K.

How many benchmarks do GLM-4.5 and MiMo-V2-Flash share?

18 benchmarks have published results for both models. GLM-4.5 has 27 scored results on Noometry and MiMo-V2-Flash has 21.

Related comparisons

Go deeper