Model comparison

Mistral Nemo vs Mistral Small

Mistral Small is the stronger model overall, scoring 33.4 to 26.4 on the Noometry Index. Mistral Nemo costs 1.7× less per token, which makes it the better buy when Mistral Small's lead doesn't matter for your workload.

Last verified . 4 shared benchmarks.

Mistral Nemo Mistral AI

26.4

Rank #337 Confirmed

Mistral Small Mistral AI

33.4

Rank #243 Confirmed

Summary

  • They share 4 benchmarks with published results for both. Mistral Nemo scores higher in 2 categories and Mistral Small in 3 categories; 4 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Mistral Small leads 52.5 to 28.5.
  • The biggest single-benchmark swing is MATH Level 5: 10.8% for Mistral Nemo and 46.8% for Mistral Small.
  • Mistral Nemo is cheaper at $0.15 / $0.15 per million input/output tokens, against $0.15 / $0.60 for Mistral Small.
  • Mistral Small accepts more context: 262K tokens versus 128K.

Side by side

Mistral Nemo and Mistral Small specifications
Mistral NemoMistral Small
ProviderMistral AIMistral AI
Noometry Index26.433.4
Released2024-07-012024-02-26
WeightsOpenOpen
Context window128K262K
Max output128K256K
Input $ / M tokens$0.15$0.15
Output $ / M tokens$0.15$0.60
Results tracked1039

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Mistral Nemo: —, Mistral Small: 34.0 (#247)

Coding benchmarks
BenchmarkMistral NemoMistral Small
SciCode—26.5%
BigCodeBench Instruct—36.1%
LiveBench Coding—36.2%
LMArena Coding—1362
BigCodeBench Complete—46.6%
ALE-Bench—497.62

Agentic & Tool Use Mistral Small leads

Mistral Nemo: 23.5 (#125), Mistral Small: 28.1 (#93)

Agentic & Tool Use benchmarks
BenchmarkMistral NemoMistral Small
Berkeley Function Calling Leaderboard27.6%37.1%
BALROG17.6%—

Reasoning Too close to call

Mistral Nemo: 20.7 (#232), Mistral Small: 19.8 (#250)

Reasoning benchmarks
BenchmarkMistral NemoMistral Small
DTBench48.6%70.9%
Kagi LLM Benchmark—37.8%
CritPt—0%
LiveBench Reasoning—44.8%
LMArena Hard Prompts—1335
LiveBench Data Analysis—53.7%
LMCA—20.6%
Epoch Capabilities Index118.68—
LiveBench—44%
PIQA83.5%—

Math Mistral Nemo leads

Mistral Nemo: 25.5 (#268), Mistral Small: 16.4 (#293)

Math benchmarks
BenchmarkMistral NemoMistral Small
MATH Level 510.8%46.8%
OTIS Mock AIME 2024-2025—5.8%
LiveBench Math—39.9%
LMArena Math—1341
GSM8K84.2%—

Knowledge Mistral Small leads

Mistral Nemo: 12.3 (#298), Mistral Small: 31.0 (#222)

Knowledge benchmarks
BenchmarkMistral NemoMistral Small
GPQA Diamond29.9%47.5%
Vectara Hallucination Rate—5.1%
LMArena Expert—1291
BoolQ82.5%—
MMLU—68.7%

Multimodal Not comparable

Mistral Nemo: —, Mistral Small: 33.5 (#96)

Multimodal benchmarks
BenchmarkMistral NemoMistral Small
LMArena Vision—1142

Multilingual Not comparable

Mistral Nemo: —, Mistral Small: 45.5 (#169)

Multilingual benchmarks
BenchmarkMistral NemoMistral Small
LMArena Non-English—1315
LMArena Chinese—1340
LMArena French—1337
LMArena German—1340
LMArena Japanese—1275
LMArena Korean—1259
LMArena Russian—1324
LMArena Spanish—1346

Instruction Following Not comparable

Mistral Nemo: —, Mistral Small: 66.4 (#209)

Instruction Following benchmarks
BenchmarkMistral NemoMistral Small
LiveBench Instruction Following—63.7%
LMArena Instruction Following—1310

Long Context Not comparable

Mistral Nemo: —, Mistral Small: 40.4 (#156)

Long Context benchmarks
BenchmarkMistral NemoMistral Small
LMArena Longer Query—1327

Writing & Preference Mistral Small leads

Mistral Nemo: 28.5 (#296), Mistral Small: 52.5 (#171)

Writing & Preference benchmarks
BenchmarkMistral NemoMistral Small
LMArena Text—1338
LMArena Creative Writing—1305
EQ-Bench Creative Writing881—
LMArena Multi-Turn—1344
LiveBench Language—30.5%

Frequently asked questions

Is Mistral Nemo better than Mistral Small?

Mistral Small is the stronger model overall, scoring 33.4 to 26.4 on the Noometry Index. Mistral Nemo costs 1.7× less per token, which makes it the better buy when Mistral Small's lead doesn't matter for your workload.

Which is cheaper, Mistral Nemo or Mistral Small?

Mistral Nemo is cheaper. It lists at $0.15 per million input tokens and $0.15 per million output tokens; Mistral Small lists at $0.15 and $0.60.

Which has the bigger context window?

Mistral Small does, with 262K tokens against 128K.

How many benchmarks do Mistral Nemo and Mistral Small share?

4 benchmarks have published results for both models. Mistral Nemo has 10 scored results on Noometry and Mistral Small has 39.

Related comparisons

Go deeper