Model comparison

Laguna M.1 vs Mistral Small 3

Laguna M.1 is the stronger model overall, scoring 32.5 to 31.2 on the Noometry Index.

Last verified . 0 shared benchmarks.

Laguna M.1 Poolside

32.5

Rank #256 Reported

Mistral Small 3 Mistral AI

31.2

Rank #278 Confirmed

Summary

  • The widest gap is in math, where Laguna M.1 leads 21.1 to 16.3.
  • Laguna M.1 accepts more context: 262K tokens versus 33K.

Side by side

Laguna M.1 and Mistral Small 3 specifications
Laguna M.1Mistral Small 3
ProviderPoolsideMistral AI
Noometry Index32.531.2
Released2026-04-282025-01-30
WeightsOpenOpen
Context window262K33K
Max output33K16K
Input $ / M tokens—$0.05
Output $ / M tokens—$0.08
Results tracked324

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Laguna M.1: 36.6 (#204), Mistral Small 3: 36.5 (#207)

Coding benchmarks
BenchmarkLaguna M.1Mistral Small 3
LMArena WebDev1349—
BigCodeBench Instruct—45.3%
LMArena Coding—1246
BigCodeBench Complete—50.4%

Reasoning Laguna M.1 leads

Laguna M.1: 23.1 (#184), Mistral Small 3: 18.9 (#273)

Reasoning benchmarks
BenchmarkLaguna M.1Mistral Small 3
Chess Puzzles—0%
LMArena Hard Prompts—1233
Surface Evolver Bench15.6%—
Epoch Capabilities Index—127.07

Math Laguna M.1 leads

Laguna M.1: 21.1 (#283), Mistral Small 3: 16.3 (#295)

Math benchmarks
BenchmarkLaguna M.1Mistral Small 3
OTIS Mock AIME 2024-2025—6.7%
ProofBench0%—
LMArena Math—1240

Knowledge Not comparable

Laguna M.1: —, Mistral Small 3: 25.1 (#263)

Knowledge benchmarks
BenchmarkLaguna M.1Mistral Small 3
GPQA Diamond—47.3%
Confabulations—25.2%
LMArena Expert—1202

Multilingual Not comparable

Laguna M.1: —, Mistral Small 3: 37.3 (#236)

Multilingual benchmarks
BenchmarkLaguna M.1Mistral Small 3
LMArena Non-English—1198
LMArena Chinese—1204
LMArena French—1203
LMArena German—1211
LMArena Japanese—1111
LMArena Korean—1188
LMArena Russian—1216

Instruction Following Not comparable

Laguna M.1: —, Mistral Small 3: 63.7 (#229)

Instruction Following benchmarks
BenchmarkLaguna M.1Mistral Small 3
LMArena Instruction Following—1214

Long Context Not comparable

Laguna M.1: —, Mistral Small 3: 37.8 (#211)

Long Context benchmarks
BenchmarkLaguna M.1Mistral Small 3
LMArena Longer Query—1246

Writing & Preference Not comparable

Laguna M.1: —, Mistral Small 3: 32.2 (#280)

Writing & Preference benchmarks
BenchmarkLaguna M.1Mistral Small 3
LMArena Text—1234
LMArena Creative Writing—1195
EQ-Bench Creative Writing—707
LMArena Multi-Turn—1217

Frequently asked questions

Is Laguna M.1 better than Mistral Small 3?

Laguna M.1 is the stronger model overall, scoring 32.5 to 31.2 on the Noometry Index.

Is Laguna M.1 or Mistral Small 3 better for coding?

They score almost the same on coding (36.6 vs 36.5); test both on your own repository before choosing.

Which has the bigger context window?

Laguna M.1 does, with 262K tokens against 33K.

How many benchmarks do Laguna M.1 and Mistral Small 3 share?

0 benchmarks have published results for both models. Laguna M.1 has 3 scored results on Noometry and Mistral Small 3 has 24.

Related comparisons

Go deeper