Model comparison

Mistral Nemo vs Mistral Small 3.2

Mistral Small 3.2 is the stronger model overall, scoring 31.2 to 26.4 on the Noometry Index.

Last verified . 3 shared benchmarks.

Mistral Nemo Mistral AI

26.4

Rank #337 Confirmed

Mistral Small 3.2 Mistral AI

31.2

Rank #280 Confirmed

Summary

  • They share 3 benchmarks with published results for both. Mistral Nemo scores higher in 1 category and Mistral Small 3.2 in 3 categories; 3 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Mistral Small 3.2 leads 45.0 to 28.5.
  • The biggest single-benchmark swing is GPQA Diamond: 29.9% for Mistral Nemo and 49.1% for Mistral Small 3.2.
  • Mistral Small 3.2 is cheaper at $0.0938 / $0.25 per million input/output tokens, against $0.15 / $0.15 for Mistral Nemo.
  • Mistral Small 3.2 accepts more context: 256K tokens versus 128K.

Side by side

Mistral Nemo and Mistral Small 3.2 specifications
Mistral NemoMistral Small 3.2
ProviderMistral AIMistral AI
Noometry Index26.431.2
Released2024-07-012025-06-20
WeightsOpenOpen
Context window128K256K
Max output128K16K
Input $ / M tokens$0.15$0.0938
Output $ / M tokens$0.15$0.25
Results tracked106

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Agentic & Tool Use Not comparable

Mistral Nemo: 23.5 (#125), Mistral Small 3.2: —

Agentic & Tool Use benchmarks
BenchmarkMistral NemoMistral Small 3.2
Berkeley Function Calling Leaderboard27.6%—
BALROG17.6%—

Reasoning Mistral Nemo leads

Mistral Nemo: 20.7 (#232), Mistral Small 3.2: 18.1 (#287)

Reasoning benchmarks
BenchmarkMistral NemoMistral Small 3.2
Epoch Capabilities Index118.68131.74
Kagi LLM Benchmark—40.4%
Chess Puzzles—1%
DTBench48.6%—
PIQA83.5%—

Math Too close to call

Mistral Nemo: 25.5 (#268), Mistral Small 3.2: 26.3 (#260)

Math benchmarks
BenchmarkMistral NemoMistral Small 3.2
OTIS Mock AIME 2024-2025—30.3%
MATH Level 510.8%—
GSM8K84.2%—

Knowledge Mistral Small 3.2 leads

Mistral Nemo: 12.3 (#298), Mistral Small 3.2: 26.7 (#256)

Knowledge benchmarks
BenchmarkMistral NemoMistral Small 3.2
GPQA Diamond29.9%49.1%
BoolQ82.5%—

Writing & Preference Mistral Small 3.2 leads

Mistral Nemo: 28.5 (#296), Mistral Small 3.2: 45.0 (#224)

Writing & Preference benchmarks
BenchmarkMistral NemoMistral Small 3.2
EQ-Bench Creative Writing8811255

Frequently asked questions

Is Mistral Nemo better than Mistral Small 3.2?

Mistral Small 3.2 is the stronger model overall, scoring 31.2 to 26.4 on the Noometry Index.

Which is cheaper, Mistral Nemo or Mistral Small 3.2?

Mistral Small 3.2 is cheaper. It lists at $0.0938 per million input tokens and $0.25 per million output tokens; Mistral Nemo lists at $0.15 and $0.15.

Which has the bigger context window?

Mistral Small 3.2 does, with 256K tokens against 128K.

How many benchmarks do Mistral Nemo and Mistral Small 3.2 share?

3 benchmarks have published results for both models. Mistral Nemo has 10 scored results on Noometry and Mistral Small 3.2 has 6.

Related comparisons

Go deeper