Model comparison

Mistral Small 3 vs Nemotron 4 340b Instruct

Nemotron 4 340b Instruct is the stronger model overall, scoring 35.9 to 31.2 on the Noometry Index.

Last verified . 16 shared benchmarks.

Mistral Small 3 Mistral AI

31.2

Rank #278 Confirmed

Nemotron 4 340b Instruct NVIDIA

35.9

Rank #223 Confirmed

Summary

  • They share 16 benchmarks with published results for both. Mistral Small 3 scores higher in 3 categories and Nemotron 4 340b Instruct in 5 categories; 5 gaps are clear of the uncertainty.
  • The widest gap is in math, where Nemotron 4 340b Instruct leads 34.3 to 16.3.

Side by side

Mistral Small 3 and Nemotron 4 340b Instruct specifications
Mistral Small 3Nemotron 4 340b Instruct
ProviderMistral AINVIDIA
Noometry Index31.235.9
Released2025-01-302024-06-14
WeightsOpenOpen
Context window33K—
Max output16K—
Input $ / M tokens$0.05—
Output $ / M tokens$0.08—
Results tracked2417

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mistral Small 3 leads

Mistral Small 3: 36.5 (#207), Nemotron 4 340b Instruct: 35.2 (#228)

Coding benchmarks
BenchmarkMistral Small 3Nemotron 4 340b Instruct
LMArena Coding12461209
BigCodeBench Instruct45.3%—
BigCodeBench Complete50.4%—

Reasoning Nemotron 4 340b Instruct leads

Mistral Small 3: 18.9 (#273), Nemotron 4 340b Instruct: 23.4 (#178)

Reasoning benchmarks
BenchmarkMistral Small 3Nemotron 4 340b Instruct
LMArena Hard Prompts12331202
Chess Puzzles0%—
Epoch Capabilities Index127.07—

Math Nemotron 4 340b Instruct leads

Mistral Small 3: 16.3 (#295), Nemotron 4 340b Instruct: 34.3 (#195)

Math benchmarks
BenchmarkMistral Small 3Nemotron 4 340b Instruct
LMArena Math12401216
OTIS Mock AIME 2024-20256.7%—

Knowledge Nemotron 4 340b Instruct leads

Mistral Small 3: 25.1 (#263), Nemotron 4 340b Instruct: 32.0 (#216)

Knowledge benchmarks
BenchmarkMistral Small 3Nemotron 4 340b Instruct
LMArena Expert12021171
GPQA Diamond47.3%—
Confabulations25.2%—

Multilingual Too close to call

Mistral Small 3: 37.3 (#236), Nemotron 4 340b Instruct: 37.9 (#233)

Multilingual benchmarks
BenchmarkMistral Small 3Nemotron 4 340b Instruct
LMArena Non-English11981206
LMArena Chinese12041214
LMArena French12031197
LMArena German12111192
LMArena Japanese11111128
LMArena Korean11881148
LMArena Russian12161219
LMArena Spanish—1195

Instruction Following Too close to call

Mistral Small 3: 63.7 (#229), Nemotron 4 340b Instruct: 63.1 (#232)

Instruction Following benchmarks
BenchmarkMistral Small 3Nemotron 4 340b Instruct
LMArena Instruction Following12141203

Long Context Too close to call

Mistral Small 3: 37.8 (#211), Nemotron 4 340b Instruct: 37.1 (#221)

Long Context benchmarks
BenchmarkMistral Small 3Nemotron 4 340b Instruct
LMArena Longer Query12461225

Writing & Preference Nemotron 4 340b Instruct leads

Mistral Small 3: 32.2 (#280), Nemotron 4 340b Instruct: 42.7 (#233)

Writing & Preference benchmarks
BenchmarkMistral Small 3Nemotron 4 340b Instruct
LMArena Text12341225
LMArena Creative Writing11951203
LMArena Multi-Turn12171214
EQ-Bench Creative Writing707—

Frequently asked questions

Is Mistral Small 3 better than Nemotron 4 340b Instruct?

Nemotron 4 340b Instruct is the stronger model overall, scoring 35.9 to 31.2 on the Noometry Index.

Is Mistral Small 3 or Nemotron 4 340b Instruct better for coding?

Mistral Small 3 scores higher on coding benchmarks: 36.5 versus 35.2 in the Noometry coding category.

How many benchmarks do Mistral Small 3 and Nemotron 4 340b Instruct share?

16 benchmarks have published results for both models. Mistral Small 3 has 24 scored results on Noometry and Nemotron 4 340b Instruct has 17.

Related comparisons

Go deeper