Model comparison

Gemini 3.1 Flash Lite vs Gemma 3 27B

Gemini 3.1 Flash Lite is the stronger model overall, scoring 40.8 to 30.8 on the Noometry Index. Gemma 3 27B costs 5.6× less per token, which makes it the better buy when Gemini 3.1 Flash Lite's lead doesn't matter for your workload.

Last verified . 28 shared benchmarks.

Gemini 3.1 Flash Lite Google

40.8

Rank #144 Confirmed

Gemma 3 27B Google

30.8

Rank #284 Confirmed

Summary

  • They share 28 benchmarks with published results for both. Gemini 3.1 Flash Lite scores higher in 10 categories and Gemma 3 27B in 0 categories; 10 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Gemini 3.1 Flash Lite leads 41.9 to 25.5.
  • The biggest single-benchmark swing is OTIS Mock AIME 2024-2025: 80% for Gemini 3.1 Flash Lite and 22.5% for Gemma 3 27B.
  • Gemma 3 27B is cheaper at $0.08 / $0.16 per million input/output tokens, against $0.25 / $1.50 for Gemini 3.1 Flash Lite.
  • Gemini 3.1 Flash Lite accepts more context: 1.05M tokens versus 131K.
  • Gemma 3 27B has downloadable open weights; the other is API-only.

Side by side

Gemini 3.1 Flash Lite and Gemma 3 27B specifications
Gemini 3.1 Flash LiteGemma 3 27B
ProviderGoogleGoogle
Noometry Index40.830.8
Released2026-03-032025-03-11
WeightsProprietaryOpen
Context window1.05M131K
Max output66K8K
Input $ / M tokens$0.25$0.08
Output $ / M tokens$1.50$0.16
Results tracked3843

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Gemini 3.1 Flash Lite leads

Gemini 3.1 Flash Lite: 37.8 (#188), Gemma 3 27B: 22.5 (#334)

Coding benchmarks
BenchmarkGemini 3.1 Flash LiteGemma 3 27B
SciCode41.9%21.2%
LMArena Coding14001322
Aider Polyglot—4.9%
LMArena WebDev1256—
WeirdML52.2%—
LiveBench Coding—39.9%
ALE-Bench797.73—

Agentic & Tool Use Gemini 3.1 Flash Lite leads

Gemini 3.1 Flash Lite: 30.2 (#79), Gemma 3 27B: 25.1 (#110)

Agentic & Tool Use benchmarks
BenchmarkGemini 3.1 Flash LiteGemma 3 27B
Berkeley Function Calling Leaderboard—29.5%
DeepResearch Bench37.3%—

Reasoning Gemini 3.1 Flash Lite leads

Gemini 3.1 Flash Lite: 22.9 (#186), Gemma 3 27B: 16.7 (#301)

Reasoning benchmarks
BenchmarkGemini 3.1 Flash LiteGemma 3 27B
Kagi LLM Benchmark67.2%40.4%
CritPt1.1%0%
Chess Puzzles25%0%
LMArena Hard Prompts14071340
DTBench76.8%52.5%
LMCA35%12.3%
Epoch Capabilities Index144.47130.04
NYT Connections (extended)8.2%—
EnigmaEval3%—
Thematic Generalization63.3%—
LiveBench Reasoning—43.8%
LiveBench Data Analysis—51.5%
ForecastBench54.4—
LiveBench—50%

Math Gemini 3.1 Flash Lite leads

Gemini 3.1 Flash Lite: 40.7 (#90), Gemma 3 27B: 25.9 (#265)

Math benchmarks
BenchmarkGemini 3.1 Flash LiteGemma 3 27B
OTIS Mock AIME 2024-202580%22.5%
LMArena Math14281312
FrontierMath (Tiers 1-3)27.7%—
LiveBench Math—55.4%
MATH Level 5—74%

Knowledge Gemini 3.1 Flash Lite leads

Gemini 3.1 Flash Lite: 41.9 (#104), Gemma 3 27B: 25.5 (#261)

Knowledge benchmarks
BenchmarkGemini 3.1 Flash LiteGemma 3 27B
GPQA Diamond81.8%47.7%
Vectara Hallucination Rate8.2%7.4%
LMArena Expert13981304
Humanity's Last Exam8.6%—
Confabulations—40.3%

Multimodal Gemini 3.1 Flash Lite leads

Gemini 3.1 Flash Lite: 39.4 (#60), Gemma 3 27B: 32.6 (#100)

Multimodal benchmarks
BenchmarkGemini 3.1 Flash LiteGemma 3 27B
LMArena Vision12401164
GeoBench—52%

Multilingual Gemini 3.1 Flash Lite leads

Gemini 3.1 Flash Lite: 52.3 (#86), Gemma 3 27B: 46.9 (#155)

Multilingual benchmarks
BenchmarkGemini 3.1 Flash LiteGemma 3 27B
LMArena Non-English14111334
LMArena Chinese14611346
LMArena French14241368
LMArena German14291362
LMArena Japanese14131287
LMArena Korean13921308
LMArena Russian14201349
LMArena Spanish14211349

Instruction Following Gemini 3.1 Flash Lite leads

Gemini 3.1 Flash Lite: 72.7 (#131), Gemma 3 27B: 70.6 (#160)

Instruction Following benchmarks
BenchmarkGemini 3.1 Flash LiteGemma 3 27B
LMArena Instruction Following13771321
LiveBench Instruction Following—74.9%

Long Context Gemini 3.1 Flash Lite leads

Gemini 3.1 Flash Lite: 42.5 (#122), Gemma 3 27B: 27.6 (#293)

Long Context benchmarks
BenchmarkGemini 3.1 Flash LiteGemma 3 27B
LMArena Longer Query13941333
Fiction.LiveBench—33.3%

Writing & Preference Gemini 3.1 Flash Lite leads

Gemini 3.1 Flash Lite: 60.9 (#94), Gemma 3 27B: 52.5 (#168)

Writing & Preference benchmarks
BenchmarkGemini 3.1 Flash LiteGemma 3 27B
LMArena Text14161358
LMArena Creative Writing14011346
LMArena Multi-Turn14171345
Short-Story Creative Writing—79.9%
EQ-Bench Creative Writing—1266
LiveBench Language—34.6%

Frequently asked questions

Is Gemini 3.1 Flash Lite better than Gemma 3 27B?

Gemini 3.1 Flash Lite is the stronger model overall, scoring 40.8 to 30.8 on the Noometry Index. Gemma 3 27B costs 5.6× less per token, which makes it the better buy when Gemini 3.1 Flash Lite's lead doesn't matter for your workload.

Which is cheaper, Gemini 3.1 Flash Lite or Gemma 3 27B?

Gemma 3 27B is cheaper. It lists at $0.08 per million input tokens and $0.16 per million output tokens; Gemini 3.1 Flash Lite lists at $0.25 and $1.50.

Is Gemini 3.1 Flash Lite or Gemma 3 27B better for coding?

Gemini 3.1 Flash Lite scores higher on coding benchmarks: 37.8 versus 22.5 in the Noometry coding category.

Which has the bigger context window?

Gemini 3.1 Flash Lite does, with 1.05M tokens against 131K.

How many benchmarks do Gemini 3.1 Flash Lite and Gemma 3 27B share?

28 benchmarks have published results for both models. Gemini 3.1 Flash Lite has 38 scored results on Noometry and Gemma 3 27B has 43.

Related comparisons

Go deeper