Model comparison

Gemini 2.5 Flash-Lite vs Gemini 2.5 Pro

Gemini 2.5 Pro is the stronger model overall, scoring 45.0 to 37.0 on the Noometry Index. Gemini 2.5 Flash-Lite costs 20× less per token, which makes it the better buy when Gemini 2.5 Pro's lead doesn't matter for your workload.

Last verified . 32 shared benchmarks.

Gemini 2.5 Flash-Lite Google

37.0

Rank #211 Confirmed

Gemini 2.5 Pro Google

45.0

Rank #75 Confirmed

Summary

  • They share 32 benchmarks with published results for both. Gemini 2.5 Flash-Lite scores higher in 1 category and Gemini 2.5 Pro in 9 categories; 10 gaps are clear of the uncertainty.
  • The widest gap is in long context, where Gemini 2.5 Pro leads 59.8 to 33.3.
  • The biggest single-benchmark swing is Fiction.LiveBench: 47.2% for Gemini 2.5 Flash-Lite and 91.7% for Gemini 2.5 Pro.
  • Gemini 2.5 Flash-Lite is cheaper at $0.10 / $0.40 per million input/output tokens, against $1.25 / $10 for Gemini 2.5 Pro.

Side by side

Gemini 2.5 Flash-Lite and Gemini 2.5 Pro specifications
Gemini 2.5 Flash-LiteGemini 2.5 Pro
ProviderGoogleGoogle
Noometry Index37.045.0
Released2025-06-172025-03-25
WeightsProprietaryProprietary
Context window1.05M1.05M
Max output66K66K
Input $ / M tokens$0.10$1.25
Output $ / M tokens$0.40$10
Results tracked3378

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Gemini 2.5 Pro leads

Gemini 2.5 Flash-Lite: 38.5 (#173), Gemini 2.5 Pro: 42.4 (#101)

Coding benchmarks
BenchmarkGemini 2.5 Flash-LiteGemini 2.5 Pro
WeirdML35.2%54%
LMArena Coding13731452
ALE-Bench325.9785.52
SWE-bench Verified—57.6%
SWE-bench Verified (bash only)—53.6%
Aider Polyglot—83.1%
LMArena WebDev—1227
SciCode—42.8%
GSO—3.9%
LiveBench Coding—85.9%
CadEval—64%
AlgoTune—1.51

Agentic & Tool Use Gemini 2.5 Pro leads

Gemini 2.5 Flash-Lite: 28.0 (#96), Gemini 2.5 Pro: 29.2 (#88)

Agentic & Tool Use benchmarks
BenchmarkGemini 2.5 Flash-LiteGemini 2.5 Pro
Terminal-Bench—32.6%
Berkeley Function Calling Leaderboard36.9%—
GDPval—23.3%
Remote Labor Index—0.8%
TheAgentCompany—30.3%
τ²-bench Banking—13.7%
DeepResearch Bench—42.8%
BALROG—43.3%
LMArena Search—1142
METR Time Horizons—55.4%
Vending-Bench 2—573.64

Reasoning Gemini 2.5 Pro leads

Gemini 2.5 Flash-Lite: 22.2 (#205), Gemini 2.5 Pro: 28.8 (#99)

Reasoning benchmarks
BenchmarkGemini 2.5 Flash-LiteGemini 2.5 Pro
Kagi LLM Benchmark40.5%70.3%
LMArena Hard Prompts13771455
DTBench62.8%82.4%
LMCA18.1%34.8%
Epoch Capabilities Index133.94145.32
ARC-AGI-2—4.9%
SimpleBench—62.4%
ARC-AGI-1—41%
CritPt—2%
Chess Puzzles—20%
EnigmaEval—5.6%
LiveBench Reasoning—89.8%
LiveBench Data Analysis—79.9%
ForecastBench—61.3
LiveBench—82.3%

Math Gemini 2.5 Flash-Lite leads

Gemini 2.5 Flash-Lite: 38.0 (#144), Gemini 2.5 Pro: 32.5 (#213)

Math benchmarks
BenchmarkGemini 2.5 Flash-LiteGemini 2.5 Pro
Omni-MATH48%41.6%
LMArena Math13731450
FrontierMath (Tiers 1-3)—24.6%
FrontierMath Tier 4—0%
OTIS Mock AIME 2024-2025—84.7%
LiveBench Math—90.2%
MATH Level 5—95.9%
FrontierMath (Feb 2025 set)—14.1%
FrontierMath Tier 4 (v1)—4.2%

Knowledge Gemini 2.5 Pro leads

Gemini 2.5 Flash-Lite: 32.5 (#210), Gemini 2.5 Pro: 56.0 (#46)

Knowledge benchmarks
BenchmarkGemini 2.5 Flash-LiteGemini 2.5 Pro
MMLU-Pro53.7%86.3%
Vectara Hallucination Rate3.3%7%
GPQA (HELM)30.9%74.9%
LMArena Expert13731452
GPQA Diamond—85.3%
Humanity's Last Exam—21.6%
Confabulations—10.6%

Multimodal Gemini 2.5 Pro leads

Gemini 2.5 Flash-Lite: 29.1 (#114), Gemini 2.5 Pro: 45.2 (#18)

Multimodal benchmarks
BenchmarkGemini 2.5 Flash-LiteGemini 2.5 Pro
LMArena Vision11981263
VPCT30%48%
GeoBench—86%
LMArena Document—1421
SpatialViz-Bench—44.7%

Multilingual Gemini 2.5 Pro leads

Gemini 2.5 Flash-Lite: 49.3 (#134), Gemini 2.5 Pro: 55.3 (#31)

Multilingual benchmarks
BenchmarkGemini 2.5 Flash-LiteGemini 2.5 Pro
LMArena Non-English13691451
LMArena Chinese14041507
LMArena French13881472
LMArena German13891487
LMArena Japanese13591461
LMArena Korean13601434
LMArena Russian13731461
LMArena Spanish13961473

Instruction Following Gemini 2.5 Pro leads

Gemini 2.5 Flash-Lite: 70.0 (#168), Gemini 2.5 Pro: 75.0 (#75)

Instruction Following benchmarks
BenchmarkGemini 2.5 Flash-LiteGemini 2.5 Pro
IFEval81%84%
LMArena Instruction Following13671437
LiveBench Instruction Following—80.6%

Long Context Gemini 2.5 Pro leads

Gemini 2.5 Flash-Lite: 33.3 (#262), Gemini 2.5 Pro: 59.8 (#5)

Long Context benchmarks
BenchmarkGemini 2.5 Flash-LiteGemini 2.5 Pro
Fiction.LiveBench47.2%91.7%
LMArena Longer Query13731449

Writing & Preference Gemini 2.5 Pro leads

Gemini 2.5 Flash-Lite: 56.8 (#135), Gemini 2.5 Pro: 63.7 (#62)

Writing & Preference benchmarks
BenchmarkGemini 2.5 Flash-LiteGemini 2.5 Pro
LMArena Text13791458
LMArena Creative Writing13671454
WildBench81.8%85.7%
LMArena Multi-Turn13661453
Short-Story Creative Writing—83.8%
EQ-Bench Creative Writing—1421
LiveBench Language—67.8%

Frequently asked questions

Is Gemini 2.5 Flash-Lite better than Gemini 2.5 Pro?

Gemini 2.5 Pro is the stronger model overall, scoring 45.0 to 37.0 on the Noometry Index. Gemini 2.5 Flash-Lite costs 20× less per token, which makes it the better buy when Gemini 2.5 Pro's lead doesn't matter for your workload.

Which is cheaper, Gemini 2.5 Flash-Lite or Gemini 2.5 Pro?

Gemini 2.5 Flash-Lite is cheaper. It lists at $0.10 per million input tokens and $0.40 per million output tokens; Gemini 2.5 Pro lists at $1.25 and $10.

Is Gemini 2.5 Flash-Lite or Gemini 2.5 Pro better for coding?

Gemini 2.5 Pro scores higher on coding benchmarks: 42.4 versus 38.5 in the Noometry coding category.

Which has the bigger context window?

Both accept 1.05M tokens.

How many benchmarks do Gemini 2.5 Flash-Lite and Gemini 2.5 Pro share?

32 benchmarks have published results for both models. Gemini 2.5 Flash-Lite has 33 scored results on Noometry and Gemini 2.5 Pro has 78.

Related comparisons

Go deeper