Model comparison

Mercury 2.5 vs Mistral

Mercury 2.5 is the stronger model overall, scoring 33.5 to 29.9 on the Noometry Index.

Last verified . 0 shared benchmarks.

Mercury 2.5 Inception

33.5

Rank #242 Reported

Mistral Mistral AI

29.9

Rank #303 Confirmed

Summary

  • The widest gap is in coding, where Mercury 2.5 leads 39.5 to 33.8.

Side by side

Mercury 2.5 and Mistral specifications
Mercury 2.5Mistral
ProviderInceptionMistral AI
Noometry Index33.529.9
Released2026-09-08—
WeightsProprietaryProprietary
Context window260K—
Max output66K—
Input $ / M tokens$0.04—
Output $ / M tokens$0.15—
Results tracked422

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mercury 2.5 leads

Mercury 2.5: 39.5 (#156), Mistral: 33.8 (#250)

Coding benchmarks
BenchmarkMercury 2.5Mistral
SciCode38.5%—
LMArena Coding—1162
ALE-Bench301.65—

Reasoning Too close to call

Mercury 2.5: 22.4 (#193), Mistral: 22.2 (#200)

Reasoning benchmarks
BenchmarkMercury 2.5Mistral
CritPt0%—
LMArena Hard Prompts—1149

Math Mercury 2.5 leads

Mercury 2.5: 23.3 (#272), Mistral: 22.3 (#278)

Math benchmarks
BenchmarkMercury 2.5Mistral
ProofBench3%—
Omni-MATH—7.2%
LMArena Math—1180

Knowledge Not comparable

Mercury 2.5: —, Mistral: 16.6 (#288)

Knowledge benchmarks
BenchmarkMercury 2.5Mistral
MMLU-Pro—27.7%
GPQA (HELM)—30.3%
LMArena Expert—1125

Multilingual Not comparable

Mercury 2.5: —, Mistral: 32.8 (#254)

Multilingual benchmarks
BenchmarkMercury 2.5Mistral
LMArena Non-English—1129
LMArena Chinese—1109
LMArena French—1180
LMArena German—1155
LMArena Japanese—1013
LMArena Korean—1032
LMArena Russian—1168
LMArena Spanish—1143

Instruction Following Not comparable

Mercury 2.5: —, Mistral: 52.6 (#288)

Instruction Following benchmarks
BenchmarkMercury 2.5Mistral
IFEval—56.8%
LMArena Instruction Following—1152

Long Context Not comparable

Mercury 2.5: —, Mistral: 35.0 (#245)

Long Context benchmarks
BenchmarkMercury 2.5Mistral
LMArena Longer Query—1153

Writing & Preference Not comparable

Mercury 2.5: —, Mistral: 37.0 (#260)

Writing & Preference benchmarks
BenchmarkMercury 2.5Mistral
LMArena Text—1165
LMArena Creative Writing—1158
WildBench—66%
LMArena Multi-Turn—1147

Frequently asked questions

Is Mercury 2.5 better than Mistral?

Mercury 2.5 is the stronger model overall, scoring 33.5 to 29.9 on the Noometry Index.

Is Mercury 2.5 or Mistral better for coding?

Mercury 2.5 scores higher on coding benchmarks: 39.5 versus 33.8 in the Noometry coding category.

How many benchmarks do Mercury 2.5 and Mistral share?

0 benchmarks have published results for both models. Mercury 2.5 has 4 scored results on Noometry and Mistral has 22.

Related comparisons

Go deeper