Model comparison

Gemini 2.0 Flash-Lite vs Gemini 2.5 Pro

Gemini 2.5 Pro is the stronger model overall, scoring 45.0 to 37.8 on the Noometry Index.

Last verified . 32 shared benchmarks.

Gemini 2.0 Flash-Lite Google

37.8

Rank #194 Confirmed

Gemini 2.5 Pro Google

45.0

Rank #75 Confirmed

Summary

  • They share 32 benchmarks with published results for both. Gemini 2.0 Flash-Lite scores higher in 1 category and Gemini 2.5 Pro in 8 categories; 9 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Gemini 2.5 Pro leads 56.0 to 35.0.
  • The biggest single-benchmark swing is LiveBench Reasoning: 50.1% for Gemini 2.0 Flash-Lite and 89.8% for Gemini 2.5 Pro.

Side by side

Gemini 2.0 Flash-Lite and Gemini 2.5 Pro specifications
Gemini 2.0 Flash-LiteGemini 2.5 Pro
ProviderGoogleGoogle
Noometry Index37.845.0
Released2025-02-052025-03-25
WeightsProprietaryProprietary
Context window—1.05M
Max output—66K
Input $ / M tokens—$1.25
Output $ / M tokens—$10
Results tracked3278

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Gemini 2.5 Pro leads

Gemini 2.0 Flash-Lite: 37.9 (#185), Gemini 2.5 Pro: 42.4 (#101)

Coding benchmarks
BenchmarkGemini 2.0 Flash-LiteGemini 2.5 Pro
LiveBench Coding47.1%85.9%
LMArena Coding13221452
SWE-bench Verified—57.6%
SWE-bench Verified (bash only)—53.6%
Aider Polyglot—83.1%
LMArena WebDev—1227
SciCode—42.8%
GSO—3.9%
WeirdML—54%
CadEval—64%
ALE-Bench—785.52
AlgoTune—1.51

Agentic & Tool Use Not comparable

Gemini 2.0 Flash-Lite: —, Gemini 2.5 Pro: 29.2 (#88)

Agentic & Tool Use benchmarks
BenchmarkGemini 2.0 Flash-LiteGemini 2.5 Pro
Terminal-Bench—32.6%
GDPval—23.3%
Remote Labor Index—0.8%
TheAgentCompany—30.3%
τ²-bench Banking—13.7%
DeepResearch Bench—42.8%
BALROG—43.3%
LMArena Search—1142
METR Time Horizons—55.4%
Vending-Bench 2—573.64

Reasoning Gemini 2.5 Pro leads

Gemini 2.0 Flash-Lite: 22.0 (#210), Gemini 2.5 Pro: 28.8 (#99)

Reasoning benchmarks
BenchmarkGemini 2.0 Flash-LiteGemini 2.5 Pro
LiveBench Reasoning50.1%89.8%
LMArena Hard Prompts13241455
DTBench52.5%82.4%
LiveBench Data Analysis65.5%79.9%
ForecastBench57.161.3
LiveBench54.3%82.3%
ARC-AGI-2—4.9%
SimpleBench—62.4%
Kagi LLM Benchmark—70.3%
ARC-AGI-1—41%
CritPt—2%
Chess Puzzles—20%
EnigmaEval—5.6%
LMCA—34.8%
Epoch Capabilities Index—145.32

Math Gemini 2.0 Flash-Lite leads

Gemini 2.0 Flash-Lite: 34.1 (#196), Gemini 2.5 Pro: 32.5 (#213)

Math benchmarks
BenchmarkGemini 2.0 Flash-LiteGemini 2.5 Pro
Omni-MATH37.4%41.6%
LiveBench Math58.1%90.2%
LMArena Math13091450
FrontierMath (Tiers 1-3)—24.6%
FrontierMath Tier 4—0%
OTIS Mock AIME 2024-2025—84.7%
MATH Level 5—95.9%
FrontierMath (Feb 2025 set)—14.1%
FrontierMath Tier 4 (v1)—4.2%

Knowledge Gemini 2.5 Pro leads

Gemini 2.0 Flash-Lite: 35.0 (#189), Gemini 2.5 Pro: 56.0 (#46)

Knowledge benchmarks
BenchmarkGemini 2.0 Flash-LiteGemini 2.5 Pro
MMLU-Pro72%86.3%
GPQA (HELM)50%74.9%
LMArena Expert13051452
GPQA Diamond—85.3%
Humanity's Last Exam—21.6%
Confabulations—10.6%
Vectara Hallucination Rate—7%

Multimodal Gemini 2.5 Pro leads

Gemini 2.0 Flash-Lite: 31.2 (#109), Gemini 2.5 Pro: 45.2 (#18)

Multimodal benchmarks
BenchmarkGemini 2.0 Flash-LiteGemini 2.5 Pro
LMArena Vision11001263
GeoBench—86%
VPCT—48%
LMArena Document—1421
SpatialViz-Bench—44.7%

Multilingual Gemini 2.5 Pro leads

Gemini 2.0 Flash-Lite: 46.0 (#161), Gemini 2.5 Pro: 55.3 (#31)

Multilingual benchmarks
BenchmarkGemini 2.0 Flash-LiteGemini 2.5 Pro
LMArena Non-English13231451
LMArena Chinese13391507
LMArena French13471472
LMArena German13061487
LMArena Japanese13011461
LMArena Korean13251434
LMArena Russian13281461
LMArena Spanish13131473

Instruction Following Gemini 2.5 Pro leads

Gemini 2.0 Flash-Lite: 70.4 (#163), Gemini 2.5 Pro: 75.0 (#75)

Instruction Following benchmarks
BenchmarkGemini 2.0 Flash-LiteGemini 2.5 Pro
LiveBench Instruction Following78.3%80.6%
IFEval82.4%84%
LMArena Instruction Following13051437

Long Context Gemini 2.5 Pro leads

Gemini 2.0 Flash-Lite: 40.1 (#160), Gemini 2.5 Pro: 59.8 (#5)

Long Context benchmarks
BenchmarkGemini 2.0 Flash-LiteGemini 2.5 Pro
LMArena Longer Query13201449
Fiction.LiveBench—91.7%

Writing & Preference Gemini 2.5 Pro leads

Gemini 2.0 Flash-Lite: 51.7 (#177), Gemini 2.5 Pro: 63.7 (#62)

Writing & Preference benchmarks
BenchmarkGemini 2.0 Flash-LiteGemini 2.5 Pro
LMArena Text13301458
LMArena Creative Writing13191454
WildBench79%85.7%
LMArena Multi-Turn13071453
LiveBench Language34.3%67.8%
Short-Story Creative Writing—83.8%
EQ-Bench Creative Writing—1421

Frequently asked questions

Is Gemini 2.0 Flash-Lite better than Gemini 2.5 Pro?

Gemini 2.5 Pro is the stronger model overall, scoring 45.0 to 37.8 on the Noometry Index.

Is Gemini 2.0 Flash-Lite or Gemini 2.5 Pro better for coding?

Gemini 2.5 Pro scores higher on coding benchmarks: 42.4 versus 37.9 in the Noometry coding category.

How many benchmarks do Gemini 2.0 Flash-Lite and Gemini 2.5 Pro share?

32 benchmarks have published results for both models. Gemini 2.0 Flash-Lite has 32 scored results on Noometry and Gemini 2.5 Pro has 78.

Related comparisons

Go deeper