Model comparison

Mistral Medium vs Mistral Nemo

Mistral Medium is the stronger model overall, scoring 36.3 to 26.4 on the Noometry Index. Mistral Nemo costs 20× less per token, which makes it the better buy when Mistral Medium's lead doesn't matter for your workload.

Last verified . 4 shared benchmarks.

Mistral Medium Mistral AI

36.3

Rank #218 Confirmed

Mistral Nemo Mistral AI

26.4

Rank #337 Confirmed

Summary

  • They share 4 benchmarks with published results for both. Mistral Medium scores higher in 5 categories and Mistral Nemo in 0 categories; 5 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Mistral Medium leads 60.0 to 28.5.
  • The biggest single-benchmark swing is MATH Level 5: 81.6% for Mistral Medium and 10.8% for Mistral Nemo.
  • Mistral Nemo is cheaper at $0.15 / $0.15 per million input/output tokens, against $1.50 / $7.50 for Mistral Medium.
  • Mistral Medium accepts more context: 262K tokens versus 128K.

Side by side

Mistral Medium and Mistral Nemo specifications
Mistral MediumMistral Nemo
ProviderMistral AIMistral AI
Noometry Index36.326.4
Released2023-12-112024-07-01
WeightsOpenOpen
Context window262K128K
Max output262K128K
Input $ / M tokens$1.50$0.15
Output $ / M tokens$7.50$0.15
Results tracked3610

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Mistral Medium: 34.2 (#243), Mistral Nemo: —

Coding benchmarks
BenchmarkMistral MediumMistral Nemo
FrontierCode8%—
SciCode40.2%—
WeirdML43.7%—
LMArena Coding1434—
ALE-Bench763.98—

Agentic & Tool Use Mistral Medium leads

Mistral Medium: 28.3 (#90), Mistral Nemo: 23.5 (#125)

Agentic & Tool Use benchmarks
BenchmarkMistral MediumMistral Nemo
Berkeley Function Calling Leaderboard37.7%27.6%
BALROG—17.6%

Reasoning Mistral Medium leads

Mistral Medium: 24.0 (#167), Mistral Nemo: 20.7 (#232)

Reasoning benchmarks
BenchmarkMistral MediumMistral Nemo
DTBench75.5%48.6%
Kagi LLM Benchmark50%—
CritPt0%—
LMArena Hard Prompts1426—
LMCA26.1%—
Surface Evolver Bench26.9%—
Epoch Capabilities Index—118.68
PIQA—83.5%

Math Mistral Medium leads

Mistral Medium: 28.1 (#245), Mistral Nemo: 25.5 (#268)

Math benchmarks
BenchmarkMistral MediumMistral Nemo
MATH Level 581.6%10.8%
OTIS Mock AIME 2024-202532.2%—
ProofBench9%—
LMArena Math1408—
FrontierMath (Feb 2025 set)0.3%—
GSM8K—84.2%

Knowledge Mistral Medium leads

Mistral Medium: 25.0 (#265), Mistral Nemo: 12.3 (#298)

Knowledge benchmarks
BenchmarkMistral MediumMistral Nemo
GPQA Diamond59.5%29.9%
Humanity's Last Exam4.5%—
Vectara Hallucination Rate22.7%—
LMArena Expert1408—
BoolQ—82.5%

Multimodal Not comparable

Mistral Medium: 35.3 (#88), Mistral Nemo: —

Multimodal benchmarks
BenchmarkMistral MediumMistral Nemo
LMArena Vision1172—

Multilingual Not comparable

Mistral Medium: 52.1 (#91), Mistral Nemo: —

Multilingual benchmarks
BenchmarkMistral MediumMistral Nemo
LMArena Non-English1408—
LMArena Chinese1447—
LMArena French1459—
LMArena German1432—
LMArena Japanese1378—
LMArena Korean1380—
LMArena Russian1411—
LMArena Spanish1433—

Instruction Following Not comparable

Mistral Medium: 73.7 (#116), Mistral Nemo: —

Instruction Following benchmarks
BenchmarkMistral MediumMistral Nemo
LMArena Instruction Following1398—

Long Context Not comparable

Mistral Medium: 42.9 (#114), Mistral Nemo: —

Long Context benchmarks
BenchmarkMistral MediumMistral Nemo
LMArena Longer Query1406—

Writing & Preference Mistral Medium leads

Mistral Medium: 60.0 (#103), Mistral Nemo: 28.5 (#296)

Writing & Preference benchmarks
BenchmarkMistral MediumMistral Nemo
LMArena Text1424—
LMArena Creative Writing1391—
Short-Story Creative Writing77.3%—
EQ-Bench Creative Writing—881
LMArena Multi-Turn1418—

Frequently asked questions

Is Mistral Medium better than Mistral Nemo?

Mistral Medium is the stronger model overall, scoring 36.3 to 26.4 on the Noometry Index. Mistral Nemo costs 20× less per token, which makes it the better buy when Mistral Medium's lead doesn't matter for your workload.

Which is cheaper, Mistral Medium or Mistral Nemo?

Mistral Nemo is cheaper. It lists at $0.15 per million input tokens and $0.15 per million output tokens; Mistral Medium lists at $1.50 and $7.50.

Which has the bigger context window?

Mistral Medium does, with 262K tokens against 128K.

How many benchmarks do Mistral Medium and Mistral Nemo share?

4 benchmarks have published results for both models. Mistral Medium has 36 scored results on Noometry and Mistral Nemo has 10.

Related comparisons

Go deeper