Model comparison

Magistral Medium vs Mistral Nemo

Magistral Medium is the stronger model overall, scoring 35.2 to 26.4 on the Noometry Index. Mistral Nemo costs 18× less per token, which makes it the better buy when Magistral Medium's lead doesn't matter for your workload.

Last verified . 0 shared benchmarks.

Magistral Medium Mistral AI

35.2

Rank #227 Confirmed

Mistral Nemo Mistral AI

26.4

Rank #337 Confirmed

Summary

  • The widest gap is in knowledge, where Magistral Medium leads 33.5 to 12.3.
  • Mistral Nemo is cheaper at $0.15 / $0.15 per million input/output tokens, against $2 / $5 for Magistral Medium.
  • Magistral Medium accepts more context: 262K tokens versus 128K.

Side by side

Magistral Medium and Mistral Nemo specifications
Magistral MediumMistral Nemo
ProviderMistral AIMistral AI
Noometry Index35.226.4
Released2025-03-172024-07-01
WeightsOpenOpen
Context window262K128K
Max output16K128K
Input $ / M tokens$2$0.15
Output $ / M tokens$5$0.15
Results tracked2210

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Magistral Medium: 39.1 (#161), Mistral Nemo: —

Coding benchmarks
BenchmarkMagistral MediumMistral Nemo
SciCode39.2%—
LMArena Coding1319—

Agentic & Tool Use Not comparable

Magistral Medium: —, Mistral Nemo: 23.5 (#125)

Agentic & Tool Use benchmarks
BenchmarkMagistral MediumMistral Nemo
Berkeley Function Calling Leaderboard—27.6%
BALROG—17.6%

Reasoning Mistral Nemo leads

Magistral Medium: 8.6 (#348), Mistral Nemo: 20.7 (#232)

Reasoning benchmarks
BenchmarkMagistral MediumMistral Nemo
ARC-AGI-20%—
Kagi LLM Benchmark16.2%—
ARC-AGI-16.1%—
CritPt0.3%—
LMArena Hard Prompts1267—
DTBench—48.6%
Epoch Capabilities Index—118.68
PIQA—83.5%

Math Magistral Medium leads

Magistral Medium: 35.1 (#189), Mistral Nemo: 25.5 (#268)

Math benchmarks
BenchmarkMagistral MediumMistral Nemo
LMArena Math1250—
MATH Level 5—10.8%
GSM8K—84.2%

Knowledge Magistral Medium leads

Magistral Medium: 33.5 (#202), Mistral Nemo: 12.3 (#298)

Knowledge benchmarks
BenchmarkMagistral MediumMistral Nemo
GPQA Diamond—29.9%
LMArena Expert1223—
BoolQ—82.5%

Multilingual Not comparable

Magistral Medium: 39.6 (#224), Mistral Nemo: —

Multilingual benchmarks
BenchmarkMagistral MediumMistral Nemo
LMArena Non-English1232—
LMArena Chinese1227—
LMArena French1267—
LMArena German1248—
LMArena Japanese1175—
LMArena Korean1125—
LMArena Russian1224—
LMArena Spanish1271—

Instruction Following Not comparable

Magistral Medium: 66.0 (#211), Mistral Nemo: —

Instruction Following benchmarks
BenchmarkMagistral MediumMistral Nemo
LMArena Instruction Following1254—

Long Context Not comparable

Magistral Medium: 39.3 (#183), Mistral Nemo: —

Long Context benchmarks
BenchmarkMagistral MediumMistral Nemo
LMArena Longer Query1295—

Writing & Preference Magistral Medium leads

Magistral Medium: 46.3 (#219), Mistral Nemo: 28.5 (#296)

Writing & Preference benchmarks
BenchmarkMagistral MediumMistral Nemo
LMArena Text1255—
LMArena Creative Writing1245—
EQ-Bench Creative Writing—881
LMArena Multi-Turn1275—

Frequently asked questions

Is Magistral Medium better than Mistral Nemo?

Magistral Medium is the stronger model overall, scoring 35.2 to 26.4 on the Noometry Index. Mistral Nemo costs 18× less per token, which makes it the better buy when Magistral Medium's lead doesn't matter for your workload.

Which is cheaper, Magistral Medium or Mistral Nemo?

Mistral Nemo is cheaper. It lists at $0.15 per million input tokens and $0.15 per million output tokens; Magistral Medium lists at $2 and $5.

Which has the bigger context window?

Magistral Medium does, with 262K tokens against 128K.

How many benchmarks do Magistral Medium and Mistral Nemo share?

0 benchmarks have published results for both models. Magistral Medium has 22 scored results on Noometry and Mistral Nemo has 10.

Related comparisons

Go deeper