Model comparison

Mercury vs Olmo 7b Instruct

Mercury is the stronger model overall, scoring 37.6 to 30.3 on the Noometry Index.

Last verified . 7 shared benchmarks.

Mercury Inception

37.6

Rank #199 Confirmed

Summary

  • They share 7 benchmarks with published results for both. Mercury scores higher in 4 categories and Olmo 7b Instruct in 1 category; 5 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Mercury leads 46.2 to 25.8.
  • Olmo 7b Instruct has downloadable open weights; the other is API-only.

Side by side

Mercury and Olmo 7b Instruct specifications
MercuryOlmo 7b Instruct
ProviderInceptionAllen Institute for AI (Ai2)
Noometry Index37.630.3
Released——
WeightsProprietaryOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked910

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mercury leads

Mercury: 38.7 (#170), Olmo 7b Instruct: 29.6 (#303)

Coding benchmarks
BenchmarkMercuryOlmo 7b Instruct
LMArena Coding13221016

Reasoning Olmo 7b Instruct leads

Mercury: 17.5 (#293), Olmo 7b Instruct: 18.8 (#274)

Reasoning benchmarks
BenchmarkMercuryOlmo 7b Instruct
LMArena Hard Prompts1285993
Kagi LLM Benchmark21.6%—

Math Not comparable

Mercury: —, Olmo 7b Instruct: 30.2 (#237)

Math benchmarks
BenchmarkMercuryOlmo 7b Instruct
LMArena Math—1018

Multilingual Mercury leads

Mercury: 41.6 (#206), Olmo 7b Instruct: 24.0 (#291)

Multilingual benchmarks
BenchmarkMercuryOlmo 7b Instruct
LMArena Non-English1260977
LMArena Chinese—1014
LMArena Russian—947

Instruction Following Mercury leads

Mercury: 65.2 (#224), Olmo 7b Instruct: 49.0 (#301)

Instruction Following benchmarks
BenchmarkMercuryOlmo 7b Instruct
LMArena Instruction Following1239978

Long Context Not comparable

Mercury: 38.4 (#198), Olmo 7b Instruct: —

Long Context benchmarks
BenchmarkMercuryOlmo 7b Instruct
LMArena Longer Query1266—

Writing & Preference Mercury leads

Mercury: 46.2 (#221), Olmo 7b Instruct: 25.8 (#303)

Writing & Preference benchmarks
BenchmarkMercuryOlmo 7b Instruct
LMArena Text12821032
LMArena Creative Writing1191990
LMArena Multi-Turn12821007

Frequently asked questions

Is Mercury better than Olmo 7b Instruct?

Mercury is the stronger model overall, scoring 37.6 to 30.3 on the Noometry Index.

Is Mercury or Olmo 7b Instruct better for coding?

Mercury scores higher on coding benchmarks: 38.7 versus 29.6 in the Noometry coding category.

How many benchmarks do Mercury and Olmo 7b Instruct share?

7 benchmarks have published results for both models. Mercury has 9 scored results on Noometry and Olmo 7b Instruct has 10.

Related comparisons

Go deeper