Model comparison

Gemini 2.5 Flash-Lite vs Gemma 2 27B

Gemini 2.5 Flash-Lite is the stronger model overall, scoring 37.0 to 29.4 on the Noometry Index.

Last verified . 20 shared benchmarks.

Gemini 2.5 Flash-Lite Google

37.0

Rank #211 Confirmed

Gemma 2 27B Google

29.4

Rank #312 Confirmed

Summary

  • They share 20 benchmarks with published results for both. Gemini 2.5 Flash-Lite scores higher in 7 categories and Gemma 2 27B in 1 category; 8 gaps are clear of the uncertainty.
  • The widest gap is in math, where Gemini 2.5 Flash-Lite leads 38.0 to 10.7.
  • The biggest single-benchmark swing is DTBench: 62.8% for Gemini 2.5 Flash-Lite and 48% for Gemma 2 27B.
  • Gemini 2.5 Flash-Lite is cheaper at $0.10 / $0.40 per million input/output tokens, against $0.65 / $0.65 for Gemma 2 27B.
  • Gemini 2.5 Flash-Lite accepts more context: 1.05M tokens versus 8K.
  • Gemma 2 27B has downloadable open weights; the other is API-only.

Side by side

Gemini 2.5 Flash-Lite and Gemma 2 27B specifications
Gemini 2.5 Flash-LiteGemma 2 27B
ProviderGoogleGoogle
Noometry Index37.029.4
Released2025-06-172024-06-24
WeightsProprietaryOpen
Context window1.05M8K
Max output66K2K
Input $ / M tokens$0.10$0.65
Output $ / M tokens$0.40$0.65
Results tracked3334

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Gemini 2.5 Flash-Lite leads

Gemini 2.5 Flash-Lite: 38.5 (#173), Gemma 2 27B: 34.1 (#246)

Coding benchmarks
BenchmarkGemini 2.5 Flash-LiteGemma 2 27B
LMArena Coding13731211
WeirdML35.2%—
BigCodeBench Instruct—42.8%
LiveBench Coding—36%
BigCodeBench Complete—52.5%
ALE-Bench325.9—

Agentic & Tool Use Not comparable

Gemini 2.5 Flash-Lite: 28.0 (#96), Gemma 2 27B: —

Agentic & Tool Use benchmarks
BenchmarkGemini 2.5 Flash-LiteGemma 2 27B
Berkeley Function Calling Leaderboard36.9%—

Reasoning Gemini 2.5 Flash-Lite leads

Gemini 2.5 Flash-Lite: 22.2 (#205), Gemma 2 27B: 15.3 (#315)

Reasoning benchmarks
BenchmarkGemini 2.5 Flash-LiteGemma 2 27B
LMArena Hard Prompts13771198
DTBench62.8%48%
LMCA18.1%7.1%
Epoch Capabilities Index133.94122.08
Kagi LLM Benchmark40.5%—
LiveBench Reasoning—28.1%
LiveBench Data Analysis—47.9%
LiveBench—38.2%

Math Gemini 2.5 Flash-Lite leads

Gemini 2.5 Flash-Lite: 38.0 (#144), Gemma 2 27B: 10.7 (#311)

Math benchmarks
BenchmarkGemini 2.5 Flash-LiteGemma 2 27B
LMArena Math13731212
OTIS Mock AIME 2024-2025—1.4%
Omni-MATH48%—
LiveBench Math—26.5%
MATH Level 5—27.9%

Knowledge Gemini 2.5 Flash-Lite leads

Gemini 2.5 Flash-Lite: 32.5 (#210), Gemma 2 27B: 19.0 (#280)

Knowledge benchmarks
BenchmarkGemini 2.5 Flash-LiteGemma 2 27B
LMArena Expert13731172
GPQA Diamond—36.5%
MMLU-Pro53.7%—
Confabulations—27.1%
Vectara Hallucination Rate3.3%—
GPQA (HELM)30.9%—
MMLU—75.7%

Multimodal Not comparable

Gemini 2.5 Flash-Lite: 29.1 (#114), Gemma 2 27B: —

Multimodal benchmarks
BenchmarkGemini 2.5 Flash-LiteGemma 2 27B
LMArena Vision1198—
VPCT30%—

Multilingual Gemini 2.5 Flash-Lite leads

Gemini 2.5 Flash-Lite: 49.3 (#134), Gemma 2 27B: 38.6 (#226)

Multilingual benchmarks
BenchmarkGemini 2.5 Flash-LiteGemma 2 27B
LMArena Non-English13691217
LMArena Chinese14041221
LMArena French13881247
LMArena German13891209
LMArena Japanese13591175
LMArena Korean13601174
LMArena Russian13731234
LMArena Spanish13961228

Instruction Following Gemini 2.5 Flash-Lite leads

Gemini 2.5 Flash-Lite: 70.0 (#168), Gemma 2 27B: 60.5 (#249)

Instruction Following benchmarks
BenchmarkGemini 2.5 Flash-LiteGemma 2 27B
LMArena Instruction Following13671206
LiveBench Instruction Following—58.1%
IFEval81%—

Long Context Gemma 2 27B leads

Gemini 2.5 Flash-Lite: 33.3 (#262), Gemma 2 27B: 37.3 (#218)

Long Context benchmarks
BenchmarkGemini 2.5 Flash-LiteGemma 2 27B
LMArena Longer Query13731231
Fiction.LiveBench47.2%—

Writing & Preference Gemini 2.5 Flash-Lite leads

Gemini 2.5 Flash-Lite: 56.8 (#135), Gemma 2 27B: 44.2 (#225)

Writing & Preference benchmarks
BenchmarkGemini 2.5 Flash-LiteGemma 2 27B
LMArena Text13791231
LMArena Creative Writing13671241
LMArena Multi-Turn13661224
WildBench81.8%—
LiveBench Language—32.6%

Frequently asked questions

Is Gemini 2.5 Flash-Lite better than Gemma 2 27B?

Gemini 2.5 Flash-Lite is the stronger model overall, scoring 37.0 to 29.4 on the Noometry Index.

Which is cheaper, Gemini 2.5 Flash-Lite or Gemma 2 27B?

Gemini 2.5 Flash-Lite is cheaper. It lists at $0.10 per million input tokens and $0.40 per million output tokens; Gemma 2 27B lists at $0.65 and $0.65.

Is Gemini 2.5 Flash-Lite or Gemma 2 27B better for coding?

Gemini 2.5 Flash-Lite scores higher on coding benchmarks: 38.5 versus 34.1 in the Noometry coding category.

Which has the bigger context window?

Gemini 2.5 Flash-Lite does, with 1.05M tokens against 8K.

How many benchmarks do Gemini 2.5 Flash-Lite and Gemma 2 27B share?

20 benchmarks have published results for both models. Gemini 2.5 Flash-Lite has 33 scored results on Noometry and Gemma 2 27B has 34.

Related comparisons

Go deeper