Model comparison

Magistral Small vs Mistral 7B

Magistral Small is the stronger model overall, scoring 30.2 to 23.0 on the Noometry Index. Mistral 7B costs 3.0× less per token, which makes it the better buy when Magistral Small's lead doesn't matter for your workload.

Last verified . 5 shared benchmarks.

Magistral Small Mistral AI

30.2

Rank #296 Confirmed

Mistral 7B Mistral AI

23.0

Rank #351 Confirmed

Summary

  • They share 5 benchmarks with published results for both. Magistral Small scores higher in 3 categories and Mistral 7B in 1 category; 4 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Magistral Small leads 30.9 to 7.4.
  • The biggest single-benchmark swing is GPQA Diamond: 56.1% for Magistral Small and 15.2% for Mistral 7B.
  • Mistral 7B is cheaper at $0.25 / $0.25 per million input/output tokens, against $0.50 / $1.50 for Magistral Small.
  • Magistral Small accepts more context: 128K tokens versus 8K.

Side by side

Magistral Small and Mistral 7B specifications
Magistral SmallMistral 7B
ProviderMistral AIMistral AI
Noometry Index30.223.0
Released2025-06-102023-09-27
WeightsOpenOpen
Context window128K8K
Max output40K8K
Input $ / M tokens$0.50$0.25
Output $ / M tokens$1.50$0.25
Results tracked1037

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Magistral Small leads

Magistral Small: 38.4 (#176), Mistral 7B: 26.4 (#326)

Coding benchmarks
BenchmarkMagistral SmallMistral 7B
SciCode35.2%—
BigCodeBench Instruct—19.5%
LMArena Coding—1082
BigCodeBench Complete—27.3%
HumanEval+—36%
MBPP+—42.1%

Reasoning Mistral 7B leads

Magistral Small: 6.8 (#350), Mistral 7B: 13.1 (#336)

Reasoning benchmarks
BenchmarkMagistral SmallMistral 7B
Chess Puzzles3%0%
DTBench61.3%42.5%
Epoch Capabilities Index133.19112.21
ARC-AGI-20%—
Kagi LLM Benchmark6.3%—
ARC-AGI-15%—
CritPt0.3%—
LMArena Hard Prompts—1067
Adversarial NLI—47.1%
BIG-Bench Hard—56.1%
HellaSwag—81%
PIQA—83%
WinoGrande—75.3%

Math Magistral Small leads

Magistral Small: 26.2 (#261), Mistral 7B: 8.1 (#325)

Math benchmarks
BenchmarkMagistral SmallMistral 7B
OTIS Mock AIME 2024-202530%0.3%
LMArena Math—1085
MATH Level 5—3.7%
GSM8K—54.4%

Knowledge Magistral Small leads

Magistral Small: 30.9 (#223), Mistral 7B: 7.4 (#311)

Knowledge benchmarks
BenchmarkMagistral SmallMistral 7B
GPQA Diamond56.1%15.2%
LMArena Expert—1036
ARC (AI2) Challenge—78.6%
BoolQ—87.4%
MMLU—62.5%
OpenBookQA—79.8%
TriviaQA—75.2%

Multilingual Not comparable

Magistral Small: —, Mistral 7B: 25.8 (#283)

Multilingual benchmarks
BenchmarkMagistral SmallMistral 7B
LMArena Non-English—1012
LMArena Chinese—1009
LMArena French—1037
LMArena German—987
LMArena Japanese—878
LMArena Russian—1018
LMArena Spanish—1026

Instruction Following Not comparable

Magistral Small: —, Mistral 7B: 54.2 (#280)

Instruction Following benchmarks
BenchmarkMagistral SmallMistral 7B
LMArena Instruction Following—1060

Long Context Not comparable

Magistral Small: —, Mistral 7B: 32.2 (#271)

Long Context benchmarks
BenchmarkMagistral SmallMistral 7B
LMArena Longer Query—1060

Writing & Preference Not comparable

Magistral Small: —, Mistral 7B: 30.7 (#286)

Writing & Preference benchmarks
BenchmarkMagistral SmallMistral 7B
LMArena Text—1090
LMArena Creative Writing—1068
LMArena Multi-Turn—1062

Frequently asked questions

Is Magistral Small better than Mistral 7B?

Magistral Small is the stronger model overall, scoring 30.2 to 23.0 on the Noometry Index. Mistral 7B costs 3.0× less per token, which makes it the better buy when Magistral Small's lead doesn't matter for your workload.

Which is cheaper, Magistral Small or Mistral 7B?

Mistral 7B is cheaper. It lists at $0.25 per million input tokens and $0.25 per million output tokens; Magistral Small lists at $0.50 and $1.50.

Is Magistral Small or Mistral 7B better for coding?

Magistral Small scores higher on coding benchmarks: 38.4 versus 26.4 in the Noometry coding category.

Which has the bigger context window?

Magistral Small does, with 128K tokens against 8K.

How many benchmarks do Magistral Small and Mistral 7B share?

5 benchmarks have published results for both models. Magistral Small has 10 scored results on Noometry and Mistral 7B has 37.

Related comparisons

Go deeper