Model comparison

Gemini 1.0 Pro vs GLM-4.5

GLM-4.5 is the stronger model overall, scoring 42.0 to 27.3 on the Noometry Index.

Last verified . 16 shared benchmarks.

Gemini 1.0 Pro Google

27.3

Rank #332 Confirmed

GLM-4.5 Z.ai (Zhipu)

42.0

Rank #122 Confirmed

Summary

  • They share 16 benchmarks with published results for both. Gemini 1.0 Pro scores higher in 0 categories and GLM-4.5 in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in math, where GLM-4.5 leads 39.0 to 9.3.
  • GLM-4.5 has downloadable open weights; the other is API-only.

Side by side

Gemini 1.0 Pro and GLM-4.5 specifications
Gemini 1.0 ProGLM-4.5
ProviderGoogleZ.ai (Zhipu)
Noometry Index27.342.0
Released2023-12-132025-07-27
WeightsProprietaryOpen
Context window—131K
Max output—98K
Input $ / M tokens—$0.60
Output $ / M tokens—$2.20
Results tracked2427

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding GLM-4.5 leads

Gemini 1.0 Pro: 32.2 (#275), GLM-4.5: 41.4 (#125)

Coding benchmarks
BenchmarkGemini 1.0 ProGLM-4.5
LMArena Coding11081434
SWE-bench Verified (bash only)—54.2%
WeirdML—40.6%
ALE-Bench—344.82
AlgoTune—1.52
HumanEval+55.5%—
MBPP+61.4%—

Reasoning GLM-4.5 leads

Gemini 1.0 Pro: 17.1 (#296), GLM-4.5: 28.6 (#100)

Reasoning benchmarks
BenchmarkGemini 1.0 ProGLM-4.5
LMArena Hard Prompts11091429
Kagi LLM Benchmark—57.9%
DTBench45.9%—
Epoch Capabilities Index117.04—

Math GLM-4.5 leads

Gemini 1.0 Pro: 9.3 (#321), GLM-4.5: 39.0 (#116)

Math benchmarks
BenchmarkGemini 1.0 ProGLM-4.5
LMArena Math11321427
OTIS Mock AIME 2024-20251.1%—
MATH Level 511.2%—

Knowledge GLM-4.5 leads

Gemini 1.0 Pro: 15.6 (#291), GLM-4.5: 35.9 (#179)

Knowledge benchmarks
BenchmarkGemini 1.0 ProGLM-4.5
LMArena Expert10591433
GPQA Diamond34%—
Humanity's Last Exam—8.3%
Confabulations—11.3%
MMLU70%—

Multilingual GLM-4.5 leads

Gemini 1.0 Pro: 33.4 (#252), GLM-4.5: 52.8 (#77)

Multilingual benchmarks
BenchmarkGemini 1.0 ProGLM-4.5
LMArena Non-English11381417
LMArena Chinese11241465
LMArena French11451418
LMArena German11251407
LMArena Japanese10231415
LMArena Russian11861414
LMArena Spanish11191454
LMArena Korean—1380

Instruction Following GLM-4.5 leads

Gemini 1.0 Pro: 57.6 (#267), GLM-4.5: 74.1 (#104)

Instruction Following benchmarks
BenchmarkGemini 1.0 ProGLM-4.5
LMArena Instruction Following11141404

Long Context GLM-4.5 leads

Gemini 1.0 Pro: 34.3 (#249), GLM-4.5: 38.2 (#201)

Long Context benchmarks
BenchmarkGemini 1.0 ProGLM-4.5
LMArena Longer Query11321412
Fiction.LiveBench—58.3%

Writing & Preference GLM-4.5 leads

Gemini 1.0 Pro: 36.0 (#264), GLM-4.5: 57.5 (#127)

Writing & Preference benchmarks
BenchmarkGemini 1.0 ProGLM-4.5
LMArena Text11491430
LMArena Creative Writing11311395
LMArena Multi-Turn11391415
Short-Story Creative Writing—73.4%
EQ-Bench Creative Writing—1343

Frequently asked questions

Is Gemini 1.0 Pro better than GLM-4.5?

GLM-4.5 is the stronger model overall, scoring 42.0 to 27.3 on the Noometry Index.

Is Gemini 1.0 Pro or GLM-4.5 better for coding?

GLM-4.5 scores higher on coding benchmarks: 41.4 versus 32.2 in the Noometry coding category.

How many benchmarks do Gemini 1.0 Pro and GLM-4.5 share?

16 benchmarks have published results for both models. Gemini 1.0 Pro has 24 scored results on Noometry and GLM-4.5 has 27.

Related comparisons

Go deeper