Model comparison

Gemini 1.0 Pro vs GLM-4.5-Air

GLM-4.5-Air is the stronger model overall, scoring 38.9 to 27.3 on the Noometry Index.

Last verified . 16 shared benchmarks.

Gemini 1.0 Pro Google

27.3

Rank #332 Confirmed

GLM-4.5-Air Z.ai (Zhipu)

38.9

Rank #177 Confirmed

Summary

  • They share 16 benchmarks with published results for both. Gemini 1.0 Pro scores higher in 0 categories and GLM-4.5-Air in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in math, where GLM-4.5-Air leads 36.2 to 9.3.
  • GLM-4.5-Air has downloadable open weights; the other is API-only.

Side by side

Gemini 1.0 Pro and GLM-4.5-Air specifications
Gemini 1.0 ProGLM-4.5-Air
ProviderGoogleZ.ai (Zhipu)
Noometry Index27.338.9
Released2023-12-132025-07-20
WeightsProprietaryOpen
Context window—131K
Max output—98K
Input $ / M tokens—$0.20
Output $ / M tokens—$1.10
Results tracked2427

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding GLM-4.5-Air leads

Gemini 1.0 Pro: 32.2 (#275), GLM-4.5-Air: 33.3 (#259)

Coding benchmarks
BenchmarkGemini 1.0 ProGLM-4.5-Air
LMArena Coding11081397
GSO—2.9%
HumanEval+55.5%—
MBPP+61.4%—

Reasoning GLM-4.5-Air leads

Gemini 1.0 Pro: 17.1 (#296), GLM-4.5-Air: 24.1 (#166)

Reasoning benchmarks
BenchmarkGemini 1.0 ProGLM-4.5-Air
LMArena Hard Prompts11091379
Kagi LLM Benchmark—43%
DTBench45.9%—
Epoch Capabilities Index117.04—
ForecastBench—59.2

Math GLM-4.5-Air leads

Gemini 1.0 Pro: 9.3 (#321), GLM-4.5-Air: 36.2 (#170)

Math benchmarks
BenchmarkGemini 1.0 ProGLM-4.5-Air
LMArena Math11321396
OTIS Mock AIME 2024-20251.1%—
Omni-MATH—39.1%
MATH Level 511.2%—

Knowledge GLM-4.5-Air leads

Gemini 1.0 Pro: 15.6 (#291), GLM-4.5-Air: 35.0 (#191)

Knowledge benchmarks
BenchmarkGemini 1.0 ProGLM-4.5-Air
LMArena Expert10591370
GPQA Diamond34%—
Humanity's Last Exam—8.1%
MMLU-Pro—76.2%
Vectara Hallucination Rate—9.3%
GPQA (HELM)—59.4%
MMLU70%—

Multilingual GLM-4.5-Air leads

Gemini 1.0 Pro: 33.4 (#252), GLM-4.5-Air: 49.1 (#135)

Multilingual benchmarks
BenchmarkGemini 1.0 ProGLM-4.5-Air
LMArena Non-English11381366
LMArena Chinese11241426
LMArena French11451399
LMArena German11251377
LMArena Japanese10231348
LMArena Russian11861373
LMArena Spanish11191386
LMArena Korean—1308

Instruction Following GLM-4.5-Air leads

Gemini 1.0 Pro: 57.6 (#267), GLM-4.5-Air: 69.6 (#171)

Instruction Following benchmarks
BenchmarkGemini 1.0 ProGLM-4.5-Air
LMArena Instruction Following11141354
IFEval—81.2%

Long Context GLM-4.5-Air leads

Gemini 1.0 Pro: 34.3 (#249), GLM-4.5-Air: 41.6 (#135)

Long Context benchmarks
BenchmarkGemini 1.0 ProGLM-4.5-Air
LMArena Longer Query11321366

Writing & Preference GLM-4.5-Air leads

Gemini 1.0 Pro: 36.0 (#264), GLM-4.5-Air: 55.9 (#139)

Writing & Preference benchmarks
BenchmarkGemini 1.0 ProGLM-4.5-Air
LMArena Text11491384
LMArena Creative Writing11311343
LMArena Multi-Turn11391371
WildBench—78.9%

Frequently asked questions

Is Gemini 1.0 Pro better than GLM-4.5-Air?

GLM-4.5-Air is the stronger model overall, scoring 38.9 to 27.3 on the Noometry Index.

Is Gemini 1.0 Pro or GLM-4.5-Air better for coding?

GLM-4.5-Air scores higher on coding benchmarks: 33.3 versus 32.2 in the Noometry coding category.

How many benchmarks do Gemini 1.0 Pro and GLM-4.5-Air share?

16 benchmarks have published results for both models. Gemini 1.0 Pro has 24 scored results on Noometry and GLM-4.5-Air has 27.

Related comparisons

Go deeper