Model comparison

Mercury 2.5 vs Mistral Nemo

Mercury 2.5 is the stronger model overall, scoring 33.5 to 26.4 on the Noometry Index.

Last verified . 0 shared benchmarks.

Mercury 2.5 Inception

33.5

Rank #242 Reported

Mistral Nemo Mistral AI

26.4

Rank #337 Confirmed

Summary

  • Mercury 2.5 is cheaper at $0.04 / $0.15 per million input/output tokens, against $0.15 / $0.15 for Mistral Nemo.
  • Mercury 2.5 accepts more context: 260K tokens versus 128K.
  • Mistral Nemo has downloadable open weights; the other is API-only.

Side by side

Mercury 2.5 and Mistral Nemo specifications
Mercury 2.5Mistral Nemo
ProviderInceptionMistral AI
Noometry Index33.526.4
Released2026-09-082024-07-01
WeightsProprietaryOpen
Context window260K128K
Max output66K128K
Input $ / M tokens$0.04$0.15
Output $ / M tokens$0.15$0.15
Results tracked410

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Mercury 2.5: 39.5 (#156), Mistral Nemo: —

Coding benchmarks
BenchmarkMercury 2.5Mistral Nemo
SciCode38.5%—
ALE-Bench301.65—

Agentic & Tool Use Not comparable

Mercury 2.5: —, Mistral Nemo: 23.5 (#125)

Agentic & Tool Use benchmarks
BenchmarkMercury 2.5Mistral Nemo
Berkeley Function Calling Leaderboard—27.6%
BALROG—17.6%

Reasoning Mercury 2.5 leads

Mercury 2.5: 22.4 (#193), Mistral Nemo: 20.7 (#232)

Reasoning benchmarks
BenchmarkMercury 2.5Mistral Nemo
CritPt0%—
DTBench—48.6%
Epoch Capabilities Index—118.68
PIQA—83.5%

Math Mistral Nemo leads

Mercury 2.5: 23.3 (#272), Mistral Nemo: 25.5 (#268)

Math benchmarks
BenchmarkMercury 2.5Mistral Nemo
ProofBench3%—
MATH Level 5—10.8%
GSM8K—84.2%

Knowledge Not comparable

Mercury 2.5: —, Mistral Nemo: 12.3 (#298)

Knowledge benchmarks
BenchmarkMercury 2.5Mistral Nemo
GPQA Diamond—29.9%
BoolQ—82.5%

Writing & Preference Not comparable

Mercury 2.5: —, Mistral Nemo: 28.5 (#296)

Writing & Preference benchmarks
BenchmarkMercury 2.5Mistral Nemo
EQ-Bench Creative Writing—881

Frequently asked questions

Is Mercury 2.5 better than Mistral Nemo?

Mercury 2.5 is the stronger model overall, scoring 33.5 to 26.4 on the Noometry Index.

Which is cheaper, Mercury 2.5 or Mistral Nemo?

Mercury 2.5 is cheaper. It lists at $0.04 per million input tokens and $0.15 per million output tokens; Mistral Nemo lists at $0.15 and $0.15.

Which has the bigger context window?

Mercury 2.5 does, with 260K tokens against 128K.

How many benchmarks do Mercury 2.5 and Mistral Nemo share?

0 benchmarks have published results for both models. Mercury 2.5 has 4 scored results on Noometry and Mistral Nemo has 10.

Related comparisons

Go deeper