Model comparison

Gemini 2.5 Flash-Lite vs GLM-4.7-Flash

GLM-4.7-Flash is the stronger model overall, scoring 38.8 to 37.0 on the Noometry Index.

Last verified . 17 shared benchmarks.

Gemini 2.5 Flash-Lite Google

37.0

Rank #211 Confirmed

GLM-4.7-Flash Z.ai (Zhipu)

38.8

Rank #180 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Gemini 2.5 Flash-Lite scores higher in 4 categories and GLM-4.7-Flash in 4 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Gemini 2.5 Flash-Lite leads 56.8 to 47.4.
  • The biggest single-benchmark swing is Vectara Hallucination Rate: 3.3% for Gemini 2.5 Flash-Lite and 9.3% for GLM-4.7-Flash.
  • GLM-4.7-Flash is cheaper at $0.06 / $0.40 per million input/output tokens, against $0.10 / $0.40 for Gemini 2.5 Flash-Lite.
  • Gemini 2.5 Flash-Lite accepts more context: 1.05M tokens versus 200K.
  • GLM-4.7-Flash has downloadable open weights; the other is API-only.

Side by side

Gemini 2.5 Flash-Lite and GLM-4.7-Flash specifications
Gemini 2.5 Flash-LiteGLM-4.7-Flash
ProviderGoogleZ.ai (Zhipu)
Noometry Index37.038.8
Released2025-06-172026-01-19
WeightsProprietaryOpen
Context window1.05M200K
Max output66K131K
Input $ / M tokens$0.10$0.06
Output $ / M tokens$0.40$0.40
Results tracked3321

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding GLM-4.7-Flash leads

Gemini 2.5 Flash-Lite: 38.5 (#173), GLM-4.7-Flash: 40.6 (#135)

Coding benchmarks
BenchmarkGemini 2.5 Flash-LiteGLM-4.7-Flash
LMArena Coding13731383
WeirdML35.2%—
ALE-Bench325.9—

Agentic & Tool Use Not comparable

Gemini 2.5 Flash-Lite: 28.0 (#96), GLM-4.7-Flash: —

Agentic & Tool Use benchmarks
BenchmarkGemini 2.5 Flash-LiteGLM-4.7-Flash
Berkeley Function Calling Leaderboard36.9%—

Reasoning Gemini 2.5 Flash-Lite leads

Gemini 2.5 Flash-Lite: 22.2 (#205), GLM-4.7-Flash: 20.9 (#229)

Reasoning benchmarks
BenchmarkGemini 2.5 Flash-LiteGLM-4.7-Flash
LMArena Hard Prompts13771356
Kagi LLM Benchmark40.5%—
Chess Puzzles—0%
DTBench62.8%—
LMCA18.1%—
Epoch Capabilities Index133.94—

Math Gemini 2.5 Flash-Lite leads

Gemini 2.5 Flash-Lite: 38.0 (#144), GLM-4.7-Flash: 36.1 (#173)

Math benchmarks
BenchmarkGemini 2.5 Flash-LiteGLM-4.7-Flash
LMArena Math13731355
OTIS Mock AIME 2024-2025—58.3%
Omni-MATH48%—

Knowledge GLM-4.7-Flash leads

Gemini 2.5 Flash-Lite: 32.5 (#210), GLM-4.7-Flash: 35.5 (#184)

Knowledge benchmarks
BenchmarkGemini 2.5 Flash-LiteGLM-4.7-Flash
Vectara Hallucination Rate3.3%9.3%
LMArena Expert13731357
GPQA Diamond—60.5%
MMLU-Pro53.7%—
GPQA (HELM)30.9%—

Multimodal Not comparable

Gemini 2.5 Flash-Lite: 29.1 (#114), GLM-4.7-Flash: —

Multimodal benchmarks
BenchmarkGemini 2.5 Flash-LiteGLM-4.7-Flash
LMArena Vision1198—
VPCT30%—

Multilingual Gemini 2.5 Flash-Lite leads

Gemini 2.5 Flash-Lite: 49.3 (#134), GLM-4.7-Flash: 46.5 (#158)

Multilingual benchmarks
BenchmarkGemini 2.5 Flash-LiteGLM-4.7-Flash
LMArena Non-English13691330
LMArena Chinese14041403
LMArena French13881332
LMArena German13891337
LMArena Korean13601283
LMArena Russian13731332
LMArena Spanish13961350
LMArena Japanese1359—

Instruction Following Too close to call

Gemini 2.5 Flash-Lite: 70.0 (#168), GLM-4.7-Flash: 70.1 (#167)

Instruction Following benchmarks
BenchmarkGemini 2.5 Flash-LiteGLM-4.7-Flash
LMArena Instruction Following13671327
IFEval81%—

Long Context GLM-4.7-Flash leads

Gemini 2.5 Flash-Lite: 33.3 (#262), GLM-4.7-Flash: 40.9 (#148)

Long Context benchmarks
BenchmarkGemini 2.5 Flash-LiteGLM-4.7-Flash
LMArena Longer Query13731345
Fiction.LiveBench47.2%—

Writing & Preference Gemini 2.5 Flash-Lite leads

Gemini 2.5 Flash-Lite: 56.8 (#135), GLM-4.7-Flash: 47.4 (#210)

Writing & Preference benchmarks
BenchmarkGemini 2.5 Flash-LiteGLM-4.7-Flash
LMArena Text13791351
LMArena Creative Writing13671297
LMArena Multi-Turn13661342
EQ-Bench Creative Writing—1125
WildBench81.8%—

Frequently asked questions

Is Gemini 2.5 Flash-Lite better than GLM-4.7-Flash?

GLM-4.7-Flash is the stronger model overall, scoring 38.8 to 37.0 on the Noometry Index.

Which is cheaper, Gemini 2.5 Flash-Lite or GLM-4.7-Flash?

GLM-4.7-Flash is cheaper. It lists at $0.06 per million input tokens and $0.40 per million output tokens; Gemini 2.5 Flash-Lite lists at $0.10 and $0.40.

Is Gemini 2.5 Flash-Lite or GLM-4.7-Flash better for coding?

GLM-4.7-Flash scores higher on coding benchmarks: 40.6 versus 38.5 in the Noometry coding category.

Which has the bigger context window?

Gemini 2.5 Flash-Lite does, with 1.05M tokens against 200K.

How many benchmarks do Gemini 2.5 Flash-Lite and GLM-4.7-Flash share?

17 benchmarks have published results for both models. Gemini 2.5 Flash-Lite has 33 scored results on Noometry and GLM-4.7-Flash has 21.

Related comparisons

Go deeper