Model comparison

Mercury vs Mistral

Mercury is the stronger model overall, scoring 37.6 to 29.9 on the Noometry Index.

Last verified . 8 shared benchmarks.

Mercury Inception

37.6

Rank #199 Confirmed

Mistral Mistral AI

29.9

Rank #303 Confirmed

Summary

  • They share 8 benchmarks with published results for both. Mercury scores higher in 5 categories and Mistral in 1 category; 6 gaps are clear of the uncertainty.
  • The widest gap is in instruction following, where Mercury leads 65.2 to 52.6.

Side by side

Mercury and Mistral specifications
MercuryMistral
ProviderInceptionMistral AI
Noometry Index37.629.9
Released——
WeightsProprietaryProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked922

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mercury leads

Mercury: 38.7 (#170), Mistral: 33.8 (#250)

Coding benchmarks
BenchmarkMercuryMistral
LMArena Coding13221162

Reasoning Mistral leads

Mercury: 17.5 (#293), Mistral: 22.2 (#200)

Reasoning benchmarks
BenchmarkMercuryMistral
LMArena Hard Prompts12851149
Kagi LLM Benchmark21.6%—

Math Not comparable

Mercury: —, Mistral: 22.3 (#278)

Math benchmarks
BenchmarkMercuryMistral
Omni-MATH—7.2%
LMArena Math—1180

Knowledge Not comparable

Mercury: —, Mistral: 16.6 (#288)

Knowledge benchmarks
BenchmarkMercuryMistral
MMLU-Pro—27.7%
GPQA (HELM)—30.3%
LMArena Expert—1125

Multilingual Mercury leads

Mercury: 41.6 (#206), Mistral: 32.8 (#254)

Multilingual benchmarks
BenchmarkMercuryMistral
LMArena Non-English12601129
LMArena Chinese—1109
LMArena French—1180
LMArena German—1155
LMArena Japanese—1013
LMArena Korean—1032
LMArena Russian—1168
LMArena Spanish—1143

Instruction Following Mercury leads

Mercury: 65.2 (#224), Mistral: 52.6 (#288)

Instruction Following benchmarks
BenchmarkMercuryMistral
LMArena Instruction Following12391152
IFEval—56.8%

Long Context Mercury leads

Mercury: 38.4 (#198), Mistral: 35.0 (#245)

Long Context benchmarks
BenchmarkMercuryMistral
LMArena Longer Query12661153

Writing & Preference Mercury leads

Mercury: 46.2 (#221), Mistral: 37.0 (#260)

Writing & Preference benchmarks
BenchmarkMercuryMistral
LMArena Text12821165
LMArena Creative Writing11911158
LMArena Multi-Turn12821147
WildBench—66%

Frequently asked questions

Is Mercury better than Mistral?

Mercury is the stronger model overall, scoring 37.6 to 29.9 on the Noometry Index.

Is Mercury or Mistral better for coding?

Mercury scores higher on coding benchmarks: 38.7 versus 33.8 in the Noometry coding category.

How many benchmarks do Mercury and Mistral share?

8 benchmarks have published results for both models. Mercury has 9 scored results on Noometry and Mistral has 22.

Related comparisons

Go deeper