Model comparison

Mercury 2.5 vs Mistral Small 3.2

Mercury 2.5 is the stronger model overall, scoring 33.5 to 31.2 on the Noometry Index.

Last verified . 0 shared benchmarks.

Mercury 2.5 Inception

33.5

Rank #242 Reported

Mistral Small 3.2 Mistral AI

31.2

Rank #280 Confirmed

Summary

  • The widest gap is in reasoning, where Mercury 2.5 leads 22.4 to 18.1.
  • Mercury 2.5 is cheaper at $0.04 / $0.15 per million input/output tokens, against $0.0938 / $0.25 for Mistral Small 3.2.
  • Mercury 2.5 accepts more context: 260K tokens versus 256K.
  • Mistral Small 3.2 has downloadable open weights; the other is API-only.

Side by side

Mercury 2.5 and Mistral Small 3.2 specifications
Mercury 2.5Mistral Small 3.2
ProviderInceptionMistral AI
Noometry Index33.531.2
Released2026-09-082025-06-20
WeightsProprietaryOpen
Context window260K256K
Max output66K16K
Input $ / M tokens$0.04$0.0938
Output $ / M tokens$0.15$0.25
Results tracked46

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Mercury 2.5: 39.5 (#156), Mistral Small 3.2: —

Coding benchmarks
BenchmarkMercury 2.5Mistral Small 3.2
SciCode38.5%—
ALE-Bench301.65—

Reasoning Mercury 2.5 leads

Mercury 2.5: 22.4 (#193), Mistral Small 3.2: 18.1 (#287)

Reasoning benchmarks
BenchmarkMercury 2.5Mistral Small 3.2
Kagi LLM Benchmark—40.4%
CritPt0%—
Chess Puzzles—1%
Epoch Capabilities Index—131.74

Math Mistral Small 3.2 leads

Mercury 2.5: 23.3 (#272), Mistral Small 3.2: 26.3 (#260)

Math benchmarks
BenchmarkMercury 2.5Mistral Small 3.2
OTIS Mock AIME 2024-2025—30.3%
ProofBench3%—

Knowledge Not comparable

Mercury 2.5: —, Mistral Small 3.2: 26.7 (#256)

Knowledge benchmarks
BenchmarkMercury 2.5Mistral Small 3.2
GPQA Diamond—49.1%

Writing & Preference Not comparable

Mercury 2.5: —, Mistral Small 3.2: 45.0 (#224)

Writing & Preference benchmarks
BenchmarkMercury 2.5Mistral Small 3.2
EQ-Bench Creative Writing—1255

Frequently asked questions

Is Mercury 2.5 better than Mistral Small 3.2?

Mercury 2.5 is the stronger model overall, scoring 33.5 to 31.2 on the Noometry Index.

Which is cheaper, Mercury 2.5 or Mistral Small 3.2?

Mercury 2.5 is cheaper. It lists at $0.04 per million input tokens and $0.15 per million output tokens; Mistral Small 3.2 lists at $0.0938 and $0.25.

Which has the bigger context window?

Mercury 2.5 does, with 260K tokens against 256K.

How many benchmarks do Mercury 2.5 and Mistral Small 3.2 share?

0 benchmarks have published results for both models. Mercury 2.5 has 4 scored results on Noometry and Mistral Small 3.2 has 6.

Related comparisons

Go deeper