Model comparison

Gemini 1.0 Pro vs Mercury 2.5

Mercury 2.5 is the stronger model overall, scoring 33.5 to 27.3 on the Noometry Index.

Last verified . 0 shared benchmarks.

Gemini 1.0 Pro Google

27.3

Rank #332 Confirmed

Mercury 2.5 Inception

33.5

Rank #242 Reported

Summary

  • The widest gap is in math, where Mercury 2.5 leads 23.3 to 9.3.

Side by side

Gemini 1.0 Pro and Mercury 2.5 specifications
Gemini 1.0 ProMercury 2.5
ProviderGoogleInception
Noometry Index27.333.5
Released2023-12-132026-09-08
WeightsProprietaryProprietary
Context window—260K
Max output—66K
Input $ / M tokens—$0.04
Output $ / M tokens—$0.15
Results tracked244

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mercury 2.5 leads

Gemini 1.0 Pro: 32.2 (#275), Mercury 2.5: 39.5 (#156)

Coding benchmarks
BenchmarkGemini 1.0 ProMercury 2.5
SciCode—38.5%
LMArena Coding1108—
ALE-Bench—301.65
HumanEval+55.5%—
MBPP+61.4%—

Reasoning Mercury 2.5 leads

Gemini 1.0 Pro: 17.1 (#296), Mercury 2.5: 22.4 (#193)

Reasoning benchmarks
BenchmarkGemini 1.0 ProMercury 2.5
CritPt—0%
LMArena Hard Prompts1109—
DTBench45.9%—
Epoch Capabilities Index117.04—

Math Mercury 2.5 leads

Gemini 1.0 Pro: 9.3 (#321), Mercury 2.5: 23.3 (#272)

Math benchmarks
BenchmarkGemini 1.0 ProMercury 2.5
OTIS Mock AIME 2024-20251.1%—
ProofBench—3%
LMArena Math1132—
MATH Level 511.2%—

Knowledge Not comparable

Gemini 1.0 Pro: 15.6 (#291), Mercury 2.5: —

Knowledge benchmarks
BenchmarkGemini 1.0 ProMercury 2.5
GPQA Diamond34%—
LMArena Expert1059—
MMLU70%—

Multilingual Not comparable

Gemini 1.0 Pro: 33.4 (#252), Mercury 2.5: —

Multilingual benchmarks
BenchmarkGemini 1.0 ProMercury 2.5
LMArena Non-English1138—
LMArena Chinese1124—
LMArena French1145—
LMArena German1125—
LMArena Japanese1023—
LMArena Russian1186—
LMArena Spanish1119—

Instruction Following Not comparable

Gemini 1.0 Pro: 57.6 (#267), Mercury 2.5: —

Instruction Following benchmarks
BenchmarkGemini 1.0 ProMercury 2.5
LMArena Instruction Following1114—

Long Context Not comparable

Gemini 1.0 Pro: 34.3 (#249), Mercury 2.5: —

Long Context benchmarks
BenchmarkGemini 1.0 ProMercury 2.5
LMArena Longer Query1132—

Writing & Preference Not comparable

Gemini 1.0 Pro: 36.0 (#264), Mercury 2.5: —

Writing & Preference benchmarks
BenchmarkGemini 1.0 ProMercury 2.5
LMArena Text1149—
LMArena Creative Writing1131—
LMArena Multi-Turn1139—

Frequently asked questions

Is Gemini 1.0 Pro better than Mercury 2.5?

Mercury 2.5 is the stronger model overall, scoring 33.5 to 27.3 on the Noometry Index.

Is Gemini 1.0 Pro or Mercury 2.5 better for coding?

Mercury 2.5 scores higher on coding benchmarks: 39.5 versus 32.2 in the Noometry coding category.

How many benchmarks do Gemini 1.0 Pro and Mercury 2.5 share?

0 benchmarks have published results for both models. Gemini 1.0 Pro has 24 scored results on Noometry and Mercury 2.5 has 4.

Related comparisons

Go deeper