Model comparison

Mercury vs o1-pro

Mercury is the stronger model overall, scoring 37.6 to 31.5 on the Noometry Index.

Last verified . 0 shared benchmarks.

Mercury Inception

37.6

Rank #199 Confirmed

o1-pro OpenAI

31.5

Rank #271 Reported

Side by side

Mercury and o1-pro specifications
Mercuryo1-pro
ProviderInceptionOpenAI
Noometry Index37.631.5
Released—2025-03-19
WeightsProprietaryProprietary
Context window—200K
Max output—100K
Input $ / M tokens—$150
Output $ / M tokens—$600
Results tracked93

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Mercury: 38.7 (#170), o1-pro: —

Coding benchmarks
BenchmarkMercuryo1-pro
LMArena Coding1322—

Reasoning o1-pro leads

Mercury: 17.5 (#293), o1-pro: 20.4 (#239)

Reasoning benchmarks
BenchmarkMercuryo1-pro
Kagi LLM Benchmark21.6%—
ARC-AGI-1—23.3%
EnigmaEval—6.1%
LMArena Hard Prompts1285—

Knowledge Not comparable

Mercury: —, o1-pro: 29.7 (#234)

Knowledge benchmarks
BenchmarkMercuryo1-pro
Humanity's Last Exam—8.1%

Multilingual Not comparable

Mercury: 41.6 (#206), o1-pro: —

Multilingual benchmarks
BenchmarkMercuryo1-pro
LMArena Non-English1260—

Instruction Following Not comparable

Mercury: 65.2 (#224), o1-pro: —

Instruction Following benchmarks
BenchmarkMercuryo1-pro
LMArena Instruction Following1239—

Long Context Not comparable

Mercury: 38.4 (#198), o1-pro: —

Long Context benchmarks
BenchmarkMercuryo1-pro
LMArena Longer Query1266—

Writing & Preference Not comparable

Mercury: 46.2 (#221), o1-pro: —

Writing & Preference benchmarks
BenchmarkMercuryo1-pro
LMArena Text1282—
LMArena Creative Writing1191—
LMArena Multi-Turn1282—

Frequently asked questions

Is Mercury better than o1-pro?

Mercury is the stronger model overall, scoring 37.6 to 31.5 on the Noometry Index.

How many benchmarks do Mercury and o1-pro share?

0 benchmarks have published results for both models. Mercury has 9 scored results on Noometry and o1-pro has 3.

Related comparisons

Go deeper