Model comparison

Mistral Small 3 vs Olmo 7b Instruct

Mistral Small 3 and Olmo 7b Instruct score almost the same on the Noometry Index (31.2 vs 30.3), so choose on price, context window or the category you care about most.

Last verified . 10 shared benchmarks.

Mistral Small 3 Mistral AI

31.2

Rank #278 Confirmed

Summary

  • They share 10 benchmarks with published results for both. Mistral Small 3 scores higher in 5 categories and Olmo 7b Instruct in 1 category; 5 gaps are clear of the uncertainty.
  • The widest gap is in instruction following, where Mistral Small 3 leads 63.7 to 49.0.

Side by side

Mistral Small 3 and Olmo 7b Instruct specifications
Mistral Small 3Olmo 7b Instruct
ProviderMistral AIAllen Institute for AI (Ai2)
Noometry Index31.230.3
Released2025-01-30—
WeightsOpenOpen
Context window33K—
Max output16K—
Input $ / M tokens$0.05—
Output $ / M tokens$0.08—
Results tracked2410

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mistral Small 3 leads

Mistral Small 3: 36.5 (#207), Olmo 7b Instruct: 29.6 (#303)

Coding benchmarks
BenchmarkMistral Small 3Olmo 7b Instruct
LMArena Coding12461016
BigCodeBench Instruct45.3%—
BigCodeBench Complete50.4%—

Reasoning Too close to call

Mistral Small 3: 18.9 (#273), Olmo 7b Instruct: 18.8 (#274)

Reasoning benchmarks
BenchmarkMistral Small 3Olmo 7b Instruct
LMArena Hard Prompts1233993
Chess Puzzles0%—
Epoch Capabilities Index127.07—

Math Olmo 7b Instruct leads

Mistral Small 3: 16.3 (#295), Olmo 7b Instruct: 30.2 (#237)

Math benchmarks
BenchmarkMistral Small 3Olmo 7b Instruct
LMArena Math12401018
OTIS Mock AIME 2024-20256.7%—

Knowledge Not comparable

Mistral Small 3: 25.1 (#263), Olmo 7b Instruct: —

Knowledge benchmarks
BenchmarkMistral Small 3Olmo 7b Instruct
GPQA Diamond47.3%—
Confabulations25.2%—
LMArena Expert1202—

Multilingual Mistral Small 3 leads

Mistral Small 3: 37.3 (#236), Olmo 7b Instruct: 24.0 (#291)

Multilingual benchmarks
BenchmarkMistral Small 3Olmo 7b Instruct
LMArena Non-English1198977
LMArena Chinese12041014
LMArena Russian1216947
LMArena French1203—
LMArena German1211—
LMArena Japanese1111—
LMArena Korean1188—

Instruction Following Mistral Small 3 leads

Mistral Small 3: 63.7 (#229), Olmo 7b Instruct: 49.0 (#301)

Instruction Following benchmarks
BenchmarkMistral Small 3Olmo 7b Instruct
LMArena Instruction Following1214978

Long Context Not comparable

Mistral Small 3: 37.8 (#211), Olmo 7b Instruct: —

Long Context benchmarks
BenchmarkMistral Small 3Olmo 7b Instruct
LMArena Longer Query1246—

Writing & Preference Mistral Small 3 leads

Mistral Small 3: 32.2 (#280), Olmo 7b Instruct: 25.8 (#303)

Writing & Preference benchmarks
BenchmarkMistral Small 3Olmo 7b Instruct
LMArena Text12341032
LMArena Creative Writing1195990
LMArena Multi-Turn12171007
EQ-Bench Creative Writing707—

Frequently asked questions

Is Mistral Small 3 better than Olmo 7b Instruct?

Mistral Small 3 and Olmo 7b Instruct score almost the same on the Noometry Index (31.2 vs 30.3), so choose on price, context window or the category you care about most.

Is Mistral Small 3 or Olmo 7b Instruct better for coding?

Mistral Small 3 scores higher on coding benchmarks: 36.5 versus 29.6 in the Noometry coding category.

How many benchmarks do Mistral Small 3 and Olmo 7b Instruct share?

10 benchmarks have published results for both models. Mistral Small 3 has 24 scored results on Noometry and Olmo 7b Instruct has 10.

Related comparisons

Go deeper