Model comparison
Gemini 1.0 Pro vs Mercury 2.5
Mercury 2.5 is the stronger model overall, scoring 33.5 to 27.3 on the Noometry Index.
Last verified . 0 shared benchmarks.
Summary
- The widest gap is in math, where Mercury 2.5 leads 23.3 to 9.3.
Side by side
| Gemini 1.0 Pro | Mercury 2.5 | |
|---|---|---|
| Provider | Inception | |
| Noometry Index | 27.3 | 33.5 |
| Released | 2023-12-13 | 2026-09-08 |
| Weights | Proprietary | Proprietary |
| Context window | — | 260K |
| Max output | — | 66K |
| Input $ / M tokens | — | $0.04 |
| Output $ / M tokens | — | $0.15 |
| Results tracked | 24 | 4 |
Sponsored placements are available on pages like this one. Advertise on Noometry
Category by category
Coding Mercury 2.5 leads
Gemini 1.0 Pro: 32.2 (#275), Mercury 2.5: 39.5 (#156)
| Benchmark | Gemini 1.0 Pro | Mercury 2.5 |
|---|---|---|
| SciCode | — | 38.5% |
| LMArena Coding | 1108 | — |
| ALE-Bench | — | 301.65 |
| HumanEval+ | 55.5% | — |
| MBPP+ | 61.4% | — |
Reasoning Mercury 2.5 leads
Gemini 1.0 Pro: 17.1 (#296), Mercury 2.5: 22.4 (#193)
| Benchmark | Gemini 1.0 Pro | Mercury 2.5 |
|---|---|---|
| CritPt | — | 0% |
| LMArena Hard Prompts | 1109 | — |
| DTBench | 45.9% | — |
| Epoch Capabilities Index | 117.04 | — |
Math Mercury 2.5 leads
Gemini 1.0 Pro: 9.3 (#321), Mercury 2.5: 23.3 (#272)
| Benchmark | Gemini 1.0 Pro | Mercury 2.5 |
|---|---|---|
| OTIS Mock AIME 2024-2025 | 1.1% | — |
| ProofBench | — | 3% |
| LMArena Math | 1132 | — |
| MATH Level 5 | 11.2% | — |
Knowledge Not comparable
Gemini 1.0 Pro: 15.6 (#291), Mercury 2.5: —
| Benchmark | Gemini 1.0 Pro | Mercury 2.5 |
|---|---|---|
| GPQA Diamond | 34% | — |
| LMArena Expert | 1059 | — |
| MMLU | 70% | — |
Multilingual Not comparable
Gemini 1.0 Pro: 33.4 (#252), Mercury 2.5: —
| Benchmark | Gemini 1.0 Pro | Mercury 2.5 |
|---|---|---|
| LMArena Non-English | 1138 | — |
| LMArena Chinese | 1124 | — |
| LMArena French | 1145 | — |
| LMArena German | 1125 | — |
| LMArena Japanese | 1023 | — |
| LMArena Russian | 1186 | — |
| LMArena Spanish | 1119 | — |
Instruction Following Not comparable
Gemini 1.0 Pro: 57.6 (#267), Mercury 2.5: —
| Benchmark | Gemini 1.0 Pro | Mercury 2.5 |
|---|---|---|
| LMArena Instruction Following | 1114 | — |
Long Context Not comparable
Gemini 1.0 Pro: 34.3 (#249), Mercury 2.5: —
| Benchmark | Gemini 1.0 Pro | Mercury 2.5 |
|---|---|---|
| LMArena Longer Query | 1132 | — |
Writing & Preference Not comparable
Gemini 1.0 Pro: 36.0 (#264), Mercury 2.5: —
| Benchmark | Gemini 1.0 Pro | Mercury 2.5 |
|---|---|---|
| LMArena Text | 1149 | — |
| LMArena Creative Writing | 1131 | — |
| LMArena Multi-Turn | 1139 | — |
Frequently asked questions
Is Gemini 1.0 Pro better than Mercury 2.5?
Mercury 2.5 is the stronger model overall, scoring 33.5 to 27.3 on the Noometry Index.
Is Gemini 1.0 Pro or Mercury 2.5 better for coding?
Mercury 2.5 scores higher on coding benchmarks: 39.5 versus 32.2 in the Noometry coding category.
How many benchmarks do Gemini 1.0 Pro and Mercury 2.5 share?
0 benchmarks have published results for both models. Gemini 1.0 Pro has 24 scored results on Noometry and Mercury 2.5 has 4.