Model comparison

Gemini 2.5 Flash-Lite vs Gemini 3.1 Flash Lite

Gemini 3.1 Flash Lite is the stronger model overall, scoring 40.8 to 37.0 on the Noometry Index. Gemini 2.5 Flash-Lite costs 3.2× less per token, which makes it the better buy when Gemini 3.1 Flash Lite's lead doesn't matter for your workload.

Last verified . 25 shared benchmarks.

Gemini 2.5 Flash-Lite Google

37.0

Rank #211 Confirmed

Gemini 3.1 Flash Lite Google

40.8

Rank #144 Confirmed

Summary

  • They share 25 benchmarks with published results for both. Gemini 2.5 Flash-Lite scores higher in 1 category and Gemini 3.1 Flash Lite in 9 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in multimodal, where Gemini 3.1 Flash Lite leads 39.4 to 29.1.
  • The biggest single-benchmark swing is Kagi LLM Benchmark: 40.5% for Gemini 2.5 Flash-Lite and 67.2% for Gemini 3.1 Flash Lite.
  • Gemini 2.5 Flash-Lite is cheaper at $0.10 / $0.40 per million input/output tokens, against $0.25 / $1.50 for Gemini 3.1 Flash Lite.

Side by side

Gemini 2.5 Flash-Lite and Gemini 3.1 Flash Lite specifications
Gemini 2.5 Flash-LiteGemini 3.1 Flash Lite
ProviderGoogleGoogle
Noometry Index37.040.8
Released2025-06-172026-03-03
WeightsProprietaryProprietary
Context window1.05M1.05M
Max output66K66K
Input $ / M tokens$0.10$0.25
Output $ / M tokens$0.40$1.50
Results tracked3338

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Gemini 2.5 Flash-Lite: 38.5 (#173), Gemini 3.1 Flash Lite: 37.8 (#188)

Coding benchmarks
BenchmarkGemini 2.5 Flash-LiteGemini 3.1 Flash Lite
WeirdML35.2%52.2%
LMArena Coding13731400
ALE-Bench325.9797.73
LMArena WebDev—1256
SciCode—41.9%

Agentic & Tool Use Gemini 3.1 Flash Lite leads

Gemini 2.5 Flash-Lite: 28.0 (#96), Gemini 3.1 Flash Lite: 30.2 (#79)

Agentic & Tool Use benchmarks
BenchmarkGemini 2.5 Flash-LiteGemini 3.1 Flash Lite
Berkeley Function Calling Leaderboard36.9%—
DeepResearch Bench—37.3%

Reasoning Too close to call

Gemini 2.5 Flash-Lite: 22.2 (#205), Gemini 3.1 Flash Lite: 22.9 (#186)

Reasoning benchmarks
BenchmarkGemini 2.5 Flash-LiteGemini 3.1 Flash Lite
Kagi LLM Benchmark40.5%67.2%
LMArena Hard Prompts13771407
DTBench62.8%76.8%
LMCA18.1%35%
Epoch Capabilities Index133.94144.47
NYT Connections (extended)—8.2%
CritPt—1.1%
Chess Puzzles—25%
EnigmaEval—3%
Thematic Generalization—63.3%
ForecastBench—54.4

Math Gemini 3.1 Flash Lite leads

Gemini 2.5 Flash-Lite: 38.0 (#144), Gemini 3.1 Flash Lite: 40.7 (#90)

Math benchmarks
BenchmarkGemini 2.5 Flash-LiteGemini 3.1 Flash Lite
LMArena Math13731428
FrontierMath (Tiers 1-3)—27.7%
OTIS Mock AIME 2024-2025—80%
Omni-MATH48%—

Knowledge Gemini 3.1 Flash Lite leads

Gemini 2.5 Flash-Lite: 32.5 (#210), Gemini 3.1 Flash Lite: 41.9 (#104)

Knowledge benchmarks
BenchmarkGemini 2.5 Flash-LiteGemini 3.1 Flash Lite
Vectara Hallucination Rate3.3%8.2%
LMArena Expert13731398
GPQA Diamond—81.8%
Humanity's Last Exam—8.6%
MMLU-Pro53.7%—
GPQA (HELM)30.9%—

Multimodal Gemini 3.1 Flash Lite leads

Gemini 2.5 Flash-Lite: 29.1 (#114), Gemini 3.1 Flash Lite: 39.4 (#60)

Multimodal benchmarks
BenchmarkGemini 2.5 Flash-LiteGemini 3.1 Flash Lite
LMArena Vision11981240
VPCT30%—

Multilingual Gemini 3.1 Flash Lite leads

Gemini 2.5 Flash-Lite: 49.3 (#134), Gemini 3.1 Flash Lite: 52.3 (#86)

Multilingual benchmarks
BenchmarkGemini 2.5 Flash-LiteGemini 3.1 Flash Lite
LMArena Non-English13691411
LMArena Chinese14041461
LMArena French13881424
LMArena German13891429
LMArena Japanese13591413
LMArena Korean13601392
LMArena Russian13731420
LMArena Spanish13961421

Instruction Following Gemini 3.1 Flash Lite leads

Gemini 2.5 Flash-Lite: 70.0 (#168), Gemini 3.1 Flash Lite: 72.7 (#131)

Instruction Following benchmarks
BenchmarkGemini 2.5 Flash-LiteGemini 3.1 Flash Lite
LMArena Instruction Following13671377
IFEval81%—

Long Context Gemini 3.1 Flash Lite leads

Gemini 2.5 Flash-Lite: 33.3 (#262), Gemini 3.1 Flash Lite: 42.5 (#122)

Long Context benchmarks
BenchmarkGemini 2.5 Flash-LiteGemini 3.1 Flash Lite
LMArena Longer Query13731394
Fiction.LiveBench47.2%—

Writing & Preference Gemini 3.1 Flash Lite leads

Gemini 2.5 Flash-Lite: 56.8 (#135), Gemini 3.1 Flash Lite: 60.9 (#94)

Writing & Preference benchmarks
BenchmarkGemini 2.5 Flash-LiteGemini 3.1 Flash Lite
LMArena Text13791416
LMArena Creative Writing13671401
LMArena Multi-Turn13661417
WildBench81.8%—

Frequently asked questions

Is Gemini 2.5 Flash-Lite better than Gemini 3.1 Flash Lite?

Gemini 3.1 Flash Lite is the stronger model overall, scoring 40.8 to 37.0 on the Noometry Index. Gemini 2.5 Flash-Lite costs 3.2× less per token, which makes it the better buy when Gemini 3.1 Flash Lite's lead doesn't matter for your workload.

Which is cheaper, Gemini 2.5 Flash-Lite or Gemini 3.1 Flash Lite?

Gemini 2.5 Flash-Lite is cheaper. It lists at $0.10 per million input tokens and $0.40 per million output tokens; Gemini 3.1 Flash Lite lists at $0.25 and $1.50.

Is Gemini 2.5 Flash-Lite or Gemini 3.1 Flash Lite better for coding?

They score almost the same on coding (38.5 vs 37.8); test both on your own repository before choosing.

Which has the bigger context window?

Both accept 1.05M tokens.

How many benchmarks do Gemini 2.5 Flash-Lite and Gemini 3.1 Flash Lite share?

25 benchmarks have published results for both models. Gemini 2.5 Flash-Lite has 33 scored results on Noometry and Gemini 3.1 Flash Lite has 38.

Related comparisons

Go deeper