Model comparison
Gemini 2.0 Flash-Lite vs Mercury
Gemini 2.0 Flash-Lite and Mercury score almost the same on the Noometry Index (37.8 vs 37.6), so choose on price, context window or the category you care about most.
Last verified . 8 shared benchmarks.
Summary
- They share 8 benchmarks with published results for both. Gemini 2.0 Flash-Lite scores higher in 5 categories and Mercury in 1 category; 5 gaps are clear of the uncertainty.
- The widest gap is in writing & preference, where Gemini 2.0 Flash-Lite leads 51.7 to 46.2.
Side by side
| Gemini 2.0 Flash-Lite | Mercury | |
|---|---|---|
| Provider | Inception | |
| Noometry Index | 37.8 | 37.6 |
| Released | 2025-02-05 | — |
| Weights | Proprietary | Proprietary |
| Context window | — | — |
| Max output | — | — |
| Input $ / M tokens | — | — |
| Output $ / M tokens | — | — |
| Results tracked | 32 | 9 |
Sponsored placements are available on pages like this one. Advertise on Noometry
Category by category
Coding Too close to call
Gemini 2.0 Flash-Lite: 37.9 (#185), Mercury: 38.7 (#170)
| Benchmark | Gemini 2.0 Flash-Lite | Mercury |
|---|---|---|
| LMArena Coding | 1322 | 1322 |
| LiveBench Coding | 47.1% | — |
Reasoning Gemini 2.0 Flash-Lite leads
Gemini 2.0 Flash-Lite: 22.0 (#210), Mercury: 17.5 (#293)
| Benchmark | Gemini 2.0 Flash-Lite | Mercury |
|---|---|---|
| LMArena Hard Prompts | 1324 | 1285 |
| Kagi LLM Benchmark | — | 21.6% |
| LiveBench Reasoning | 50.1% | — |
| DTBench | 52.5% | — |
| LiveBench Data Analysis | 65.5% | — |
| ForecastBench | 57.1 | — |
| LiveBench | 54.3% | — |
Math Not comparable
Gemini 2.0 Flash-Lite: 34.1 (#196), Mercury: —
| Benchmark | Gemini 2.0 Flash-Lite | Mercury |
|---|---|---|
| Omni-MATH | 37.4% | — |
| LiveBench Math | 58.1% | — |
| LMArena Math | 1309 | — |
Knowledge Not comparable
Gemini 2.0 Flash-Lite: 35.0 (#189), Mercury: —
| Benchmark | Gemini 2.0 Flash-Lite | Mercury |
|---|---|---|
| MMLU-Pro | 72% | — |
| GPQA (HELM) | 50% | — |
| LMArena Expert | 1305 | — |
Multimodal Not comparable
Gemini 2.0 Flash-Lite: 31.2 (#109), Mercury: —
| Benchmark | Gemini 2.0 Flash-Lite | Mercury |
|---|---|---|
| LMArena Vision | 1100 | — |
Multilingual Gemini 2.0 Flash-Lite leads
Gemini 2.0 Flash-Lite: 46.0 (#161), Mercury: 41.6 (#206)
| Benchmark | Gemini 2.0 Flash-Lite | Mercury |
|---|---|---|
| LMArena Non-English | 1323 | 1260 |
| LMArena Chinese | 1339 | — |
| LMArena French | 1347 | — |
| LMArena German | 1306 | — |
| LMArena Japanese | 1301 | — |
| LMArena Korean | 1325 | — |
| LMArena Russian | 1328 | — |
| LMArena Spanish | 1313 | — |
Instruction Following Gemini 2.0 Flash-Lite leads
Gemini 2.0 Flash-Lite: 70.4 (#163), Mercury: 65.2 (#224)
| Benchmark | Gemini 2.0 Flash-Lite | Mercury |
|---|---|---|
| LMArena Instruction Following | 1305 | 1239 |
| LiveBench Instruction Following | 78.3% | — |
| IFEval | 82.4% | — |
Long Context Gemini 2.0 Flash-Lite leads
Gemini 2.0 Flash-Lite: 40.1 (#160), Mercury: 38.4 (#198)
| Benchmark | Gemini 2.0 Flash-Lite | Mercury |
|---|---|---|
| LMArena Longer Query | 1320 | 1266 |
Writing & Preference Gemini 2.0 Flash-Lite leads
Gemini 2.0 Flash-Lite: 51.7 (#177), Mercury: 46.2 (#221)
| Benchmark | Gemini 2.0 Flash-Lite | Mercury |
|---|---|---|
| LMArena Text | 1330 | 1282 |
| LMArena Creative Writing | 1319 | 1191 |
| LMArena Multi-Turn | 1307 | 1282 |
| WildBench | 79% | — |
| LiveBench Language | 34.3% | — |
Frequently asked questions
Is Gemini 2.0 Flash-Lite better than Mercury?
Gemini 2.0 Flash-Lite and Mercury score almost the same on the Noometry Index (37.8 vs 37.6), so choose on price, context window or the category you care about most.
Is Gemini 2.0 Flash-Lite or Mercury better for coding?
They score almost the same on coding (37.9 vs 38.7); test both on your own repository before choosing.
How many benchmarks do Gemini 2.0 Flash-Lite and Mercury share?
8 benchmarks have published results for both models. Gemini 2.0 Flash-Lite has 32 scored results on Noometry and Mercury has 9.