Model comparison

Mercury vs Mistral Medium 3.5

Mistral Medium 3.5 is the stronger model overall, scoring 40.2 to 37.6 on the Noometry Index.

Last verified . 9 shared benchmarks.

Mercury Inception

37.6

Rank #199 Confirmed

Mistral Medium 3.5 Mistral AI

40.2

Rank #152 Confirmed

Summary

  • They share 9 benchmarks with published results for both. Mercury scores higher in 2 categories and Mistral Medium 3.5 in 4 categories; 5 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Mistral Medium 3.5 leads 58.5 to 46.2.
  • The biggest single-benchmark swing is Kagi LLM Benchmark: 21.6% for Mercury and 41.4% for Mistral Medium 3.5.
  • Mistral Medium 3.5 has downloadable open weights; the other is API-only.

Side by side

Mercury and Mistral Medium 3.5 specifications
MercuryMistral Medium 3.5
ProviderInceptionMistral AI
Noometry Index37.640.2
Released——
WeightsProprietaryOpen
Context window—262K
Max output—210K
Input $ / M tokens—$1.50
Output $ / M tokens—$7.50
Results tracked922

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mercury leads

Mercury: 38.7 (#170), Mistral Medium 3.5: 36.0 (#213)

Coding benchmarks
BenchmarkMercuryMistral Medium 3.5
LMArena Coding13221461
LMArena WebDev—1264

Reasoning Too close to call

Mercury: 17.5 (#293), Mistral Medium 3.5: 17.3 (#295)

Reasoning benchmarks
BenchmarkMercuryMistral Medium 3.5
Kagi LLM Benchmark21.6%41.4%
LMArena Hard Prompts12851436
NYT Connections (extended)—12.9%
Epoch Capabilities Index—141.35

Math Not comparable

Mercury: —, Mistral Medium 3.5: 39.1 (#113)

Math benchmarks
BenchmarkMercuryMistral Medium 3.5
LMArena Math—1431

Knowledge Not comparable

Mercury: —, Mistral Medium 3.5: 40.0 (#126)

Knowledge benchmarks
BenchmarkMercuryMistral Medium 3.5
LMArena Expert—1432

Multimodal Not comparable

Mercury: —, Mistral Medium 3.5: 38.3 (#65)

Multimodal benchmarks
BenchmarkMercuryMistral Medium 3.5
LMArena Vision—1223

Multilingual Mistral Medium 3.5 leads

Mercury: 41.6 (#206), Mistral Medium 3.5: 51.9 (#100)

Multilingual benchmarks
BenchmarkMercuryMistral Medium 3.5
LMArena Non-English12601404
LMArena Chinese—1442
LMArena French—1448
LMArena German—1451
LMArena Korean—1385
LMArena Russian—1395
LMArena Spanish—1409

Instruction Following Mistral Medium 3.5 leads

Mercury: 65.2 (#224), Mistral Medium 3.5: 74.6 (#90)

Instruction Following benchmarks
BenchmarkMercuryMistral Medium 3.5
LMArena Instruction Following12391415

Long Context Mistral Medium 3.5 leads

Mercury: 38.4 (#198), Mistral Medium 3.5: 43.2 (#103)

Long Context benchmarks
BenchmarkMercuryMistral Medium 3.5
LMArena Longer Query12661415

Writing & Preference Mistral Medium 3.5 leads

Mercury: 46.2 (#221), Mistral Medium 3.5: 58.5 (#117)

Writing & Preference benchmarks
BenchmarkMercuryMistral Medium 3.5
LMArena Text12821421
LMArena Creative Writing11911374
LMArena Multi-Turn12821423
EQ-Bench 4—993

Frequently asked questions

Is Mercury better than Mistral Medium 3.5?

Mistral Medium 3.5 is the stronger model overall, scoring 40.2 to 37.6 on the Noometry Index.

Is Mercury or Mistral Medium 3.5 better for coding?

Mercury scores higher on coding benchmarks: 38.7 versus 36.0 in the Noometry coding category.

How many benchmarks do Mercury and Mistral Medium 3.5 share?

9 benchmarks have published results for both models. Mercury has 9 scored results on Noometry and Mistral Medium 3.5 has 22.

Related comparisons

Go deeper