Model comparison

Magistral Small vs Mistral Nemo

Magistral Small is the stronger model overall, scoring 30.2 to 26.4 on the Noometry Index. Mistral Nemo costs 5.0× less per token, which makes it the better buy when Magistral Small's lead doesn't matter for your workload.

Last verified . 3 shared benchmarks.

Magistral Small Mistral AI

30.2

Rank #296 Confirmed

Mistral Nemo Mistral AI

26.4

Rank #337 Confirmed

Summary

  • They share 3 benchmarks with published results for both. Magistral Small scores higher in 2 categories and Mistral Nemo in 1 category; 2 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Magistral Small leads 30.9 to 12.3.
  • The biggest single-benchmark swing is GPQA Diamond: 56.1% for Magistral Small and 29.9% for Mistral Nemo.
  • Mistral Nemo is cheaper at $0.15 / $0.15 per million input/output tokens, against $0.50 / $1.50 for Magistral Small.

Side by side

Magistral Small and Mistral Nemo specifications
Magistral SmallMistral Nemo
ProviderMistral AIMistral AI
Noometry Index30.226.4
Released2025-06-102024-07-01
WeightsOpenOpen
Context window128K128K
Max output40K128K
Input $ / M tokens$0.50$0.15
Output $ / M tokens$1.50$0.15
Results tracked1010

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Magistral Small: 38.4 (#176), Mistral Nemo: —

Coding benchmarks
BenchmarkMagistral SmallMistral Nemo
SciCode35.2%—

Agentic & Tool Use Not comparable

Magistral Small: —, Mistral Nemo: 23.5 (#125)

Agentic & Tool Use benchmarks
BenchmarkMagistral SmallMistral Nemo
Berkeley Function Calling Leaderboard—27.6%
BALROG—17.6%

Reasoning Mistral Nemo leads

Magistral Small: 6.8 (#350), Mistral Nemo: 20.7 (#232)

Reasoning benchmarks
BenchmarkMagistral SmallMistral Nemo
DTBench61.3%48.6%
Epoch Capabilities Index133.19118.68
ARC-AGI-20%—
Kagi LLM Benchmark6.3%—
ARC-AGI-15%—
CritPt0.3%—
Chess Puzzles3%—
PIQA—83.5%

Math Too close to call

Magistral Small: 26.2 (#261), Mistral Nemo: 25.5 (#268)

Math benchmarks
BenchmarkMagistral SmallMistral Nemo
OTIS Mock AIME 2024-202530%—
MATH Level 5—10.8%
GSM8K—84.2%

Knowledge Magistral Small leads

Magistral Small: 30.9 (#223), Mistral Nemo: 12.3 (#298)

Knowledge benchmarks
BenchmarkMagistral SmallMistral Nemo
GPQA Diamond56.1%29.9%
BoolQ—82.5%

Writing & Preference Not comparable

Magistral Small: —, Mistral Nemo: 28.5 (#296)

Writing & Preference benchmarks
BenchmarkMagistral SmallMistral Nemo
EQ-Bench Creative Writing—881

Frequently asked questions

Is Magistral Small better than Mistral Nemo?

Magistral Small is the stronger model overall, scoring 30.2 to 26.4 on the Noometry Index. Mistral Nemo costs 5.0× less per token, which makes it the better buy when Magistral Small's lead doesn't matter for your workload.

Which is cheaper, Magistral Small or Mistral Nemo?

Mistral Nemo is cheaper. It lists at $0.15 per million input tokens and $0.15 per million output tokens; Magistral Small lists at $0.50 and $1.50.

Which has the bigger context window?

Both accept 128K tokens.

How many benchmarks do Magistral Small and Mistral Nemo share?

3 benchmarks have published results for both models. Magistral Small has 10 scored results on Noometry and Mistral Nemo has 10.

Related comparisons

Go deeper