Model comparison
Magistral Small vs Mistral Small 3.2
Magistral Small and Mistral Small 3.2 score almost the same on the Noometry Index (30.2 vs 31.2), so choose on price, context window or the category you care about most.
Last verified . 5 shared benchmarks.
Summary
- They share 5 benchmarks with published results for both. Magistral Small scores higher in 1 category and Mistral Small 3.2 in 2 categories; 2 gaps are clear of the uncertainty.
- The widest gap is in reasoning, where Mistral Small 3.2 leads 18.1 to 6.8.
- The biggest single-benchmark swing is Kagi LLM Benchmark: 6.3% for Magistral Small and 40.4% for Mistral Small 3.2.
- Mistral Small 3.2 is cheaper at $0.0938 / $0.25 per million input/output tokens, against $0.50 / $1.50 for Magistral Small.
- Mistral Small 3.2 accepts more context: 256K tokens versus 128K.
Side by side
| Magistral Small | Mistral Small 3.2 | |
|---|---|---|
| Provider | Mistral AI | Mistral AI |
| Noometry Index | 30.2 | 31.2 |
| Released | 2025-06-10 | 2025-06-20 |
| Weights | Open | Open |
| Context window | 128K | 256K |
| Max output | 40K | 16K |
| Input $ / M tokens | $0.50 | $0.0938 |
| Output $ / M tokens | $1.50 | $0.25 |
| Results tracked | 10 | 6 |
Sponsored placements are available on pages like this one. Advertise on Noometry
Category by category
Coding Not comparable
Magistral Small: 38.4 (#176), Mistral Small 3.2: —
| Benchmark | Magistral Small | Mistral Small 3.2 |
|---|---|---|
| SciCode | 35.2% | — |
Reasoning Mistral Small 3.2 leads
Magistral Small: 6.8 (#350), Mistral Small 3.2: 18.1 (#287)
| Benchmark | Magistral Small | Mistral Small 3.2 |
|---|---|---|
| Kagi LLM Benchmark | 6.3% | 40.4% |
| Chess Puzzles | 3% | 1% |
| Epoch Capabilities Index | 133.19 | 131.74 |
| ARC-AGI-2 | 0% | — |
| ARC-AGI-1 | 5% | — |
| CritPt | 0.3% | — |
| DTBench | 61.3% | — |
Math Too close to call
Magistral Small: 26.2 (#261), Mistral Small 3.2: 26.3 (#260)
| Benchmark | Magistral Small | Mistral Small 3.2 |
|---|---|---|
| OTIS Mock AIME 2024-2025 | 30% | 30.3% |
Knowledge Magistral Small leads
Magistral Small: 30.9 (#223), Mistral Small 3.2: 26.7 (#256)
| Benchmark | Magistral Small | Mistral Small 3.2 |
|---|---|---|
| GPQA Diamond | 56.1% | 49.1% |
Writing & Preference Not comparable
Magistral Small: —, Mistral Small 3.2: 45.0 (#224)
| Benchmark | Magistral Small | Mistral Small 3.2 |
|---|---|---|
| EQ-Bench Creative Writing | — | 1255 |
Frequently asked questions
Is Magistral Small better than Mistral Small 3.2?
Magistral Small and Mistral Small 3.2 score almost the same on the Noometry Index (30.2 vs 31.2), so choose on price, context window or the category you care about most.
Which is cheaper, Magistral Small or Mistral Small 3.2?
Mistral Small 3.2 is cheaper. It lists at $0.0938 per million input tokens and $0.25 per million output tokens; Magistral Small lists at $0.50 and $1.50.
Which has the bigger context window?
Mistral Small 3.2 does, with 256K tokens against 128K.
How many benchmarks do Magistral Small and Mistral Small 3.2 share?
5 benchmarks have published results for both models. Magistral Small has 10 scored results on Noometry and Mistral Small 3.2 has 6.