Model comparison

Gemini 2.5 Flash-Lite vs GLM-4.6

GLM-4.6 is the stronger model overall, scoring 41.4 to 37.0 on the Noometry Index. Gemini 2.5 Flash-Lite costs 5.7× less per token, which makes it the better buy when GLM-4.6's lead doesn't matter for your workload.

Last verified . 21 shared benchmarks.

Gemini 2.5 Flash-Lite Google

37.0

Rank #211 Confirmed

GLM-4.6 Z.ai (Zhipu)

41.4

Rank #135 Confirmed

Summary

  • They share 21 benchmarks with published results for both. Gemini 2.5 Flash-Lite scores higher in 0 categories and GLM-4.6 in 9 categories; 9 gaps are clear of the uncertainty.
  • The widest gap is in long context, where GLM-4.6 leads 43.4 to 33.3.
  • The biggest single-benchmark swing is Berkeley Function Calling Leaderboard: 36.9% for Gemini 2.5 Flash-Lite and 72.4% for GLM-4.6.
  • Gemini 2.5 Flash-Lite is cheaper at $0.10 / $0.40 per million input/output tokens, against $0.60 / $2.20 for GLM-4.6.
  • Gemini 2.5 Flash-Lite accepts more context: 1.05M tokens versus 205K.
  • GLM-4.6 has downloadable open weights; the other is API-only.

Side by side

Gemini 2.5 Flash-Lite and GLM-4.6 specifications
Gemini 2.5 Flash-LiteGLM-4.6
ProviderGoogleZ.ai (Zhipu)
Noometry Index37.041.4
Released2025-06-172025-09-30
WeightsProprietaryOpen
Context window1.05M205K
Max output66K131K
Input $ / M tokens$0.10$0.60
Output $ / M tokens$0.40$2.20
Results tracked3329

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding GLM-4.6 leads

Gemini 2.5 Flash-Lite: 38.5 (#173), GLM-4.6: 40.1 (#148)

Coding benchmarks
BenchmarkGemini 2.5 Flash-LiteGLM-4.6
LMArena Coding13731449
ALE-Bench325.9340.82
SWE-bench Verified (bash only)—55.4%
LMArena WebDev—1340
SciCode—38.4%
WeirdML35.2%—

Agentic & Tool Use GLM-4.6 leads

Gemini 2.5 Flash-Lite: 28.0 (#96), GLM-4.6: 32.3 (#66)

Agentic & Tool Use benchmarks
BenchmarkGemini 2.5 Flash-LiteGLM-4.6
Berkeley Function Calling Leaderboard36.9%72.4%
Terminal-Bench—24.5%

Reasoning GLM-4.6 leads

Gemini 2.5 Flash-Lite: 22.2 (#205), GLM-4.6: 23.7 (#172)

Reasoning benchmarks
BenchmarkGemini 2.5 Flash-LiteGLM-4.6
Kagi LLM Benchmark40.5%47.4%
LMArena Hard Prompts13771440
CritPt—1.1%
DTBench62.8%—
LMCA18.1%—
Epoch Capabilities Index133.94—

Math GLM-4.6 leads

Gemini 2.5 Flash-Lite: 38.0 (#144), GLM-4.6: 39.1 (#111)

Math benchmarks
BenchmarkGemini 2.5 Flash-LiteGLM-4.6
LMArena Math13731432
Omni-MATH48%—
FrontierMath (Feb 2025 set)—3.8%
FrontierMath Tier 4 (v1)—2.1%

Knowledge GLM-4.6 leads

Gemini 2.5 Flash-Lite: 32.5 (#210), GLM-4.6: 40.2 (#124)

Knowledge benchmarks
BenchmarkGemini 2.5 Flash-LiteGLM-4.6
Vectara Hallucination Rate3.3%9.5%
LMArena Expert13731431
MMLU-Pro53.7%—
GPQA (HELM)30.9%—

Multimodal Not comparable

Gemini 2.5 Flash-Lite: 29.1 (#114), GLM-4.6: —

Multimodal benchmarks
BenchmarkGemini 2.5 Flash-LiteGLM-4.6
LMArena Vision1198—
VPCT30%—

Multilingual GLM-4.6 leads

Gemini 2.5 Flash-Lite: 49.3 (#134), GLM-4.6: 53.5 (#66)

Multilingual benchmarks
BenchmarkGemini 2.5 Flash-LiteGLM-4.6
LMArena Non-English13691426
LMArena Chinese14041499
LMArena French13881459
LMArena German13891447
LMArena Japanese13591393
LMArena Korean13601400
LMArena Russian13731419
LMArena Spanish13961436

Instruction Following GLM-4.6 leads

Gemini 2.5 Flash-Lite: 70.0 (#168), GLM-4.6: 74.3 (#98)

Instruction Following benchmarks
BenchmarkGemini 2.5 Flash-LiteGLM-4.6
LMArena Instruction Following13671410
IFEval81%—

Long Context GLM-4.6 leads

Gemini 2.5 Flash-Lite: 33.3 (#262), GLM-4.6: 43.4 (#94)

Long Context benchmarks
BenchmarkGemini 2.5 Flash-LiteGLM-4.6
LMArena Longer Query13731422
Fiction.LiveBench47.2%—

Writing & Preference GLM-4.6 leads

Gemini 2.5 Flash-Lite: 56.8 (#135), GLM-4.6: 61.1 (#90)

Writing & Preference benchmarks
BenchmarkGemini 2.5 Flash-LiteGLM-4.6
LMArena Text13791440
LMArena Creative Writing13671411
LMArena Multi-Turn13661427
EQ-Bench Creative Writing—1411
WildBench81.8%—

Frequently asked questions

Is Gemini 2.5 Flash-Lite better than GLM-4.6?

GLM-4.6 is the stronger model overall, scoring 41.4 to 37.0 on the Noometry Index. Gemini 2.5 Flash-Lite costs 5.7× less per token, which makes it the better buy when GLM-4.6's lead doesn't matter for your workload.

Which is cheaper, Gemini 2.5 Flash-Lite or GLM-4.6?

Gemini 2.5 Flash-Lite is cheaper. It lists at $0.10 per million input tokens and $0.40 per million output tokens; GLM-4.6 lists at $0.60 and $2.20.

Is Gemini 2.5 Flash-Lite or GLM-4.6 better for coding?

GLM-4.6 scores higher on coding benchmarks: 40.1 versus 38.5 in the Noometry coding category.

Which has the bigger context window?

Gemini 2.5 Flash-Lite does, with 1.05M tokens against 205K.

How many benchmarks do Gemini 2.5 Flash-Lite and GLM-4.6 share?

21 benchmarks have published results for both models. Gemini 2.5 Flash-Lite has 33 scored results on Noometry and GLM-4.6 has 29.

Related comparisons

Go deeper