Model comparison

Gemini 2.0 Flash-Lite vs Mercury

Gemini 2.0 Flash-Lite and Mercury score almost the same on the Noometry Index (37.8 vs 37.6), so choose on price, context window or the category you care about most.

Last verified . 8 shared benchmarks.

Gemini 2.0 Flash-Lite Google

37.8

Rank #194 Confirmed

Mercury Inception

37.6

Rank #199 Confirmed

Summary

  • They share 8 benchmarks with published results for both. Gemini 2.0 Flash-Lite scores higher in 5 categories and Mercury in 1 category; 5 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Gemini 2.0 Flash-Lite leads 51.7 to 46.2.

Side by side

Gemini 2.0 Flash-Lite and Mercury specifications
Gemini 2.0 Flash-LiteMercury
ProviderGoogleInception
Noometry Index37.837.6
Released2025-02-05—
WeightsProprietaryProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked329

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Gemini 2.0 Flash-Lite: 37.9 (#185), Mercury: 38.7 (#170)

Coding benchmarks
BenchmarkGemini 2.0 Flash-LiteMercury
LMArena Coding13221322
LiveBench Coding47.1%—

Reasoning Gemini 2.0 Flash-Lite leads

Gemini 2.0 Flash-Lite: 22.0 (#210), Mercury: 17.5 (#293)

Reasoning benchmarks
BenchmarkGemini 2.0 Flash-LiteMercury
LMArena Hard Prompts13241285
Kagi LLM Benchmark—21.6%
LiveBench Reasoning50.1%—
DTBench52.5%—
LiveBench Data Analysis65.5%—
ForecastBench57.1—
LiveBench54.3%—

Math Not comparable

Gemini 2.0 Flash-Lite: 34.1 (#196), Mercury: —

Math benchmarks
BenchmarkGemini 2.0 Flash-LiteMercury
Omni-MATH37.4%—
LiveBench Math58.1%—
LMArena Math1309—

Knowledge Not comparable

Gemini 2.0 Flash-Lite: 35.0 (#189), Mercury: —

Knowledge benchmarks
BenchmarkGemini 2.0 Flash-LiteMercury
MMLU-Pro72%—
GPQA (HELM)50%—
LMArena Expert1305—

Multimodal Not comparable

Gemini 2.0 Flash-Lite: 31.2 (#109), Mercury: —

Multimodal benchmarks
BenchmarkGemini 2.0 Flash-LiteMercury
LMArena Vision1100—

Multilingual Gemini 2.0 Flash-Lite leads

Gemini 2.0 Flash-Lite: 46.0 (#161), Mercury: 41.6 (#206)

Multilingual benchmarks
BenchmarkGemini 2.0 Flash-LiteMercury
LMArena Non-English13231260
LMArena Chinese1339—
LMArena French1347—
LMArena German1306—
LMArena Japanese1301—
LMArena Korean1325—
LMArena Russian1328—
LMArena Spanish1313—

Instruction Following Gemini 2.0 Flash-Lite leads

Gemini 2.0 Flash-Lite: 70.4 (#163), Mercury: 65.2 (#224)

Instruction Following benchmarks
BenchmarkGemini 2.0 Flash-LiteMercury
LMArena Instruction Following13051239
LiveBench Instruction Following78.3%—
IFEval82.4%—

Long Context Gemini 2.0 Flash-Lite leads

Gemini 2.0 Flash-Lite: 40.1 (#160), Mercury: 38.4 (#198)

Long Context benchmarks
BenchmarkGemini 2.0 Flash-LiteMercury
LMArena Longer Query13201266

Writing & Preference Gemini 2.0 Flash-Lite leads

Gemini 2.0 Flash-Lite: 51.7 (#177), Mercury: 46.2 (#221)

Writing & Preference benchmarks
BenchmarkGemini 2.0 Flash-LiteMercury
LMArena Text13301282
LMArena Creative Writing13191191
LMArena Multi-Turn13071282
WildBench79%—
LiveBench Language34.3%—

Frequently asked questions

Is Gemini 2.0 Flash-Lite better than Mercury?

Gemini 2.0 Flash-Lite and Mercury score almost the same on the Noometry Index (37.8 vs 37.6), so choose on price, context window or the category you care about most.

Is Gemini 2.0 Flash-Lite or Mercury better for coding?

They score almost the same on coding (37.9 vs 38.7); test both on your own repository before choosing.

How many benchmarks do Gemini 2.0 Flash-Lite and Mercury share?

8 benchmarks have published results for both models. Gemini 2.0 Flash-Lite has 32 scored results on Noometry and Mercury has 9.

Related comparisons

Go deeper