Model comparison

Mistral Medium 3.5 vs Mistral Small 3.2

Mistral Medium 3.5 is the stronger model overall, scoring 40.2 to 31.2 on the Noometry Index. Mistral Small 3.2 costs 23× less per token, which makes it the better buy when Mistral Medium 3.5's lead doesn't matter for your workload.

Last verified . 2 shared benchmarks.

Mistral Medium 3.5 Mistral AI

40.2

Rank #152 Confirmed

Mistral Small 3.2 Mistral AI

31.2

Rank #280 Confirmed

Summary

  • They share 2 benchmarks with published results for both. Mistral Medium 3.5 scores higher in 3 categories and Mistral Small 3.2 in 1 category; 3 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Mistral Medium 3.5 leads 58.5 to 45.0.
  • Mistral Small 3.2 is cheaper at $0.0938 / $0.25 per million input/output tokens, against $1.50 / $7.50 for Mistral Medium 3.5.
  • Mistral Medium 3.5 accepts more context: 262K tokens versus 256K.

Side by side

Mistral Medium 3.5 and Mistral Small 3.2 specifications
Mistral Medium 3.5Mistral Small 3.2
ProviderMistral AIMistral AI
Noometry Index40.231.2
Released—2025-06-20
WeightsOpenOpen
Context window262K256K
Max output210K16K
Input $ / M tokens$1.50$0.0938
Output $ / M tokens$7.50$0.25
Results tracked226

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Mistral Medium 3.5: 36.0 (#213), Mistral Small 3.2: —

Coding benchmarks
BenchmarkMistral Medium 3.5Mistral Small 3.2
LMArena WebDev1264—
LMArena Coding1461—

Reasoning Too close to call

Mistral Medium 3.5: 17.3 (#295), Mistral Small 3.2: 18.1 (#287)

Reasoning benchmarks
BenchmarkMistral Medium 3.5Mistral Small 3.2
Kagi LLM Benchmark41.4%40.4%
Epoch Capabilities Index141.35131.74
NYT Connections (extended)12.9%—
Chess Puzzles—1%
LMArena Hard Prompts1436—

Math Mistral Medium 3.5 leads

Mistral Medium 3.5: 39.1 (#113), Mistral Small 3.2: 26.3 (#260)

Math benchmarks
BenchmarkMistral Medium 3.5Mistral Small 3.2
OTIS Mock AIME 2024-2025—30.3%
LMArena Math1431—

Knowledge Mistral Medium 3.5 leads

Mistral Medium 3.5: 40.0 (#126), Mistral Small 3.2: 26.7 (#256)

Knowledge benchmarks
BenchmarkMistral Medium 3.5Mistral Small 3.2
GPQA Diamond—49.1%
LMArena Expert1432—

Multimodal Not comparable

Mistral Medium 3.5: 38.3 (#65), Mistral Small 3.2: —

Multimodal benchmarks
BenchmarkMistral Medium 3.5Mistral Small 3.2
LMArena Vision1223—

Multilingual Not comparable

Mistral Medium 3.5: 51.9 (#100), Mistral Small 3.2: —

Multilingual benchmarks
BenchmarkMistral Medium 3.5Mistral Small 3.2
LMArena Non-English1404—
LMArena Chinese1442—
LMArena French1448—
LMArena German1451—
LMArena Korean1385—
LMArena Russian1395—
LMArena Spanish1409—

Instruction Following Not comparable

Mistral Medium 3.5: 74.6 (#90), Mistral Small 3.2: —

Instruction Following benchmarks
BenchmarkMistral Medium 3.5Mistral Small 3.2
LMArena Instruction Following1415—

Long Context Not comparable

Mistral Medium 3.5: 43.2 (#103), Mistral Small 3.2: —

Long Context benchmarks
BenchmarkMistral Medium 3.5Mistral Small 3.2
LMArena Longer Query1415—

Writing & Preference Mistral Medium 3.5 leads

Mistral Medium 3.5: 58.5 (#117), Mistral Small 3.2: 45.0 (#224)

Writing & Preference benchmarks
BenchmarkMistral Medium 3.5Mistral Small 3.2
LMArena Text1421—
LMArena Creative Writing1374—
EQ-Bench Creative Writing—1255
EQ-Bench 4993—
LMArena Multi-Turn1423—

Frequently asked questions

Is Mistral Medium 3.5 better than Mistral Small 3.2?

Mistral Medium 3.5 is the stronger model overall, scoring 40.2 to 31.2 on the Noometry Index. Mistral Small 3.2 costs 23× less per token, which makes it the better buy when Mistral Medium 3.5's lead doesn't matter for your workload.

Which is cheaper, Mistral Medium 3.5 or Mistral Small 3.2?

Mistral Small 3.2 is cheaper. It lists at $0.0938 per million input tokens and $0.25 per million output tokens; Mistral Medium 3.5 lists at $1.50 and $7.50.

Which has the bigger context window?

Mistral Medium 3.5 does, with 262K tokens against 256K.

How many benchmarks do Mistral Medium 3.5 and Mistral Small 3.2 share?

2 benchmarks have published results for both models. Mistral Medium 3.5 has 22 scored results on Noometry and Mistral Small 3.2 has 6.

Related comparisons

Go deeper