Model comparison

Mercury vs Ministral 8B

Mercury is the stronger model overall, scoring 37.6 to 28.2 on the Noometry Index.

Last verified . 8 shared benchmarks.

Mercury Inception

37.6

Rank #199 Confirmed

Ministral 8B Mistral AI

28.2

Rank #325 Confirmed

Summary

  • They share 8 benchmarks with published results for both. Mercury scores higher in 5 categories and Ministral 8B in 1 category; 5 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Mercury leads 46.2 to 39.6.
  • Ministral 8B has downloadable open weights; the other is API-only.

Side by side

Mercury and Ministral 8B specifications
MercuryMinistral 8B
ProviderInceptionMistral AI
Noometry Index37.628.2
Released—2024-10-01
WeightsProprietaryOpen
Context window—262K
Max output—262K
Input $ / M tokens—$0.15
Output $ / M tokens—$0.15
Results tracked917

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mercury leads

Mercury: 38.7 (#170), Ministral 8B: 35.0 (#230)

Coding benchmarks
BenchmarkMercuryMinistral 8B
LMArena Coding13221202

Agentic & Tool Use Not comparable

Mercury: —, Ministral 8B: 16.4 (#148)

Agentic & Tool Use benchmarks
BenchmarkMercuryMinistral 8B
Berkeley Function Calling Leaderboard—11.1%

Reasoning Too close to call

Mercury: 17.5 (#293), Ministral 8B: 18.4 (#281)

Reasoning benchmarks
BenchmarkMercuryMinistral 8B
LMArena Hard Prompts12851191
Kagi LLM Benchmark21.6%—
DTBench—45.7%

Math Not comparable

Mercury: —, Ministral 8B: 25.7 (#267)

Math benchmarks
BenchmarkMercuryMinistral 8B
LMArena Math—1188
MATH Level 5—14.9%

Knowledge Not comparable

Mercury: —, Ministral 8B: 12.6 (#297)

Knowledge benchmarks
BenchmarkMercuryMinistral 8B
GPQA Diamond—27.1%
Vectara Hallucination Rate—7.4%
LMArena Expert—1170

Multilingual Mercury leads

Mercury: 41.6 (#206), Ministral 8B: 35.1 (#247)

Multilingual benchmarks
BenchmarkMercuryMinistral 8B
LMArena Non-English12601165
LMArena Chinese—1193
LMArena Russian—1195

Instruction Following Mercury leads

Mercury: 65.2 (#224), Ministral 8B: 60.5 (#250)

Instruction Following benchmarks
BenchmarkMercuryMinistral 8B
LMArena Instruction Following12391161

Long Context Mercury leads

Mercury: 38.4 (#198), Ministral 8B: 36.7 (#227)

Long Context benchmarks
BenchmarkMercuryMinistral 8B
LMArena Longer Query12661212

Writing & Preference Mercury leads

Mercury: 46.2 (#221), Ministral 8B: 39.6 (#246)

Writing & Preference benchmarks
BenchmarkMercuryMinistral 8B
LMArena Text12821191
LMArena Creative Writing11911175
LMArena Multi-Turn12821166

Frequently asked questions

Is Mercury better than Ministral 8B?

Mercury is the stronger model overall, scoring 37.6 to 28.2 on the Noometry Index.

Is Mercury or Ministral 8B better for coding?

Mercury scores higher on coding benchmarks: 38.7 versus 35.0 in the Noometry coding category.

How many benchmarks do Mercury and Ministral 8B share?

8 benchmarks have published results for both models. Mercury has 9 scored results on Noometry and Ministral 8B has 17.

Related comparisons

Go deeper