Model comparison

GLM-4.7 vs MiniMax M1

GLM-4.7 is the stronger model overall, scoring 42.0 to 40.3 on the Noometry Index.

Last verified . 17 shared benchmarks.

GLM-4.7 Z.ai (Zhipu)

42.0

Rank #124 Confirmed

MiniMax M1 MiniMax

40.3

Rank #150 Confirmed

Summary

  • They share 17 benchmarks with published results for both. GLM-4.7 scores higher in 7 categories and MiniMax M1 in 1 category; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where GLM-4.7 leads 47.0 to 36.4.
  • Both cost about the same: $0.60 input and $2.20 output per million tokens.
  • MiniMax M1 accepts more context: 1M tokens versus 205K.

Side by side

GLM-4.7 and MiniMax M1 specifications
GLM-4.7MiniMax M1
ProviderZ.ai (Zhipu)MiniMax
Noometry Index42.040.3
Released2025-12-222025-06-13
WeightsOpenOpen
Context window205K1M
Max output131K40K
Input $ / M tokens$0.60$0.55
Output $ / M tokens$2.20$2.20
Results tracked3618

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding GLM-4.7 leads

GLM-4.7: 44.0 (#79), MiniMax M1: 39.9 (#153)

Coding benchmarks
BenchmarkGLM-4.7MiniMax M1
LMArena Coding14541359
LMArena WebDev1435—
SciCode45.1%—
ALE-Bench399.48—

Agentic & Tool Use Not comparable

GLM-4.7: 26.5 (#103), MiniMax M1: —

Agentic & Tool Use benchmarks
BenchmarkGLM-4.7MiniMax M1
Terminal-Bench33.4%—
Vending-Bench 22,377—

Reasoning MiniMax M1 leads

GLM-4.7: 24.3 (#164), MiniMax M1: 26.9 (#126)

Reasoning benchmarks
BenchmarkGLM-4.7MiniMax M1
LMArena Hard Prompts14431339
SimpleBench47.7%—
CritPt1.7%—
Chess Puzzles6%—
Epoch Capabilities Index143.51—

Math GLM-4.7 leads

GLM-4.7: 38.6 (#135), MiniMax M1: 37.5 (#151)

Math benchmarks
BenchmarkGLM-4.7MiniMax M1
LMArena Math14231361
OTIS Mock AIME 2024-202583.3%—
ProofBench6%—
FrontierMath (Feb 2025 set)2.4%—
FrontierMath Tier 4 (v1)0%—

Knowledge GLM-4.7 leads

GLM-4.7: 47.0 (#80), MiniMax M1: 36.4 (#170)

Knowledge benchmarks
BenchmarkGLM-4.7MiniMax M1
LMArena Expert14241317
GPQA Diamond83.3%—
SimpleQA Verified32.2%—
Vectara Hallucination Rate11.7%—

Multilingual GLM-4.7 leads

GLM-4.7: 52.8 (#79), MiniMax M1: 45.8 (#163)

Multilingual benchmarks
BenchmarkGLM-4.7MiniMax M1
LMArena Non-English14171319
LMArena Chinese14951360
LMArena French14321370
LMArena German14241350
LMArena Japanese14391217
LMArena Korean13991266
LMArena Russian14231329
LMArena Spanish14341353

Instruction Following GLM-4.7 leads

GLM-4.7: 74.4 (#95), MiniMax M1: 69.3 (#174)

Instruction Following benchmarks
BenchmarkGLM-4.7MiniMax M1
LMArena Instruction Following14111312

Long Context GLM-4.7 leads

GLM-4.7: 42.8 (#116), MiniMax M1: 41.4 (#141)

Long Context benchmarks
BenchmarkGLM-4.7MiniMax M1
LMArena Longer Query14321326
Fiction.LiveBench—69.4%
CL-bench15.9%—
CL-bench Life10.9%—

Writing & Preference GLM-4.7 leads

GLM-4.7: 60.9 (#93), MiniMax M1: 53.1 (#161)

Writing & Preference benchmarks
BenchmarkGLM-4.7MiniMax M1
LMArena Text14351343
LMArena Creative Writing14011298
LMArena Multi-Turn14461335
EQ-Bench Creative Writing1413—

Frequently asked questions

Is GLM-4.7 better than MiniMax M1?

GLM-4.7 is the stronger model overall, scoring 42.0 to 40.3 on the Noometry Index.

Which is cheaper, GLM-4.7 or MiniMax M1?

MiniMax M1 is cheaper. It lists at $0.55 per million input tokens and $2.20 per million output tokens; GLM-4.7 lists at $0.60 and $2.20.

Is GLM-4.7 or MiniMax M1 better for coding?

GLM-4.7 scores higher on coding benchmarks: 44.0 versus 39.9 in the Noometry coding category.

Which has the bigger context window?

MiniMax M1 does, with 1M tokens against 205K.

How many benchmarks do GLM-4.7 and MiniMax M1 share?

17 benchmarks have published results for both models. GLM-4.7 has 36 scored results on Noometry and MiniMax M1 has 18.

Related comparisons

Go deeper