Model comparison

Mistral vs Mistral 7B

Mistral is the stronger model overall, scoring 29.9 to 23.0 on the Noometry Index.

Last verified . 16 shared benchmarks.

Mistral Mistral AI

29.9

Rank #303 Confirmed

Mistral 7B Mistral AI

23.0

Rank #351 Confirmed

Summary

  • They share 16 benchmarks with published results for both. Mistral scores higher in 7 categories and Mistral 7B in 1 category; 8 gaps are clear of the uncertainty.
  • The widest gap is in math, where Mistral leads 22.3 to 8.1.
  • Mistral 7B has downloadable open weights; the other is API-only.

Side by side

Mistral and Mistral 7B specifications
MistralMistral 7B
ProviderMistral AIMistral AI
Noometry Index29.923.0
Released—2023-09-27
WeightsProprietaryOpen
Context window—8K
Max output—8K
Input $ / M tokens—$0.25
Output $ / M tokens—$0.25
Results tracked2237

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mistral leads

Mistral: 33.8 (#250), Mistral 7B: 26.4 (#326)

Coding benchmarks
BenchmarkMistralMistral 7B
LMArena Coding11621082
BigCodeBench Instruct—19.5%
BigCodeBench Complete—27.3%
HumanEval+—36%
MBPP+—42.1%

Reasoning Mistral leads

Mistral: 22.2 (#200), Mistral 7B: 13.1 (#336)

Reasoning benchmarks
BenchmarkMistralMistral 7B
LMArena Hard Prompts11491067
Chess Puzzles—0%
DTBench—42.5%
Adversarial NLI—47.1%
BIG-Bench Hard—56.1%
Epoch Capabilities Index—112.21
HellaSwag—81%
PIQA—83%
WinoGrande—75.3%

Math Mistral leads

Mistral: 22.3 (#278), Mistral 7B: 8.1 (#325)

Math benchmarks
BenchmarkMistralMistral 7B
LMArena Math11801085
OTIS Mock AIME 2024-2025—0.3%
Omni-MATH7.2%—
MATH Level 5—3.7%
GSM8K—54.4%

Knowledge Mistral leads

Mistral: 16.6 (#288), Mistral 7B: 7.4 (#311)

Knowledge benchmarks
BenchmarkMistralMistral 7B
LMArena Expert11251036
GPQA Diamond—15.2%
MMLU-Pro27.7%—
GPQA (HELM)30.3%—
ARC (AI2) Challenge—78.6%
BoolQ—87.4%
MMLU—62.5%
OpenBookQA—79.8%
TriviaQA—75.2%

Multilingual Mistral leads

Mistral: 32.8 (#254), Mistral 7B: 25.8 (#283)

Multilingual benchmarks
BenchmarkMistralMistral 7B
LMArena Non-English11291012
LMArena Chinese11091009
LMArena French11801037
LMArena German1155987
LMArena Japanese1013878
LMArena Russian11681018
LMArena Spanish11431026
LMArena Korean1032—

Instruction Following Mistral 7B leads

Mistral: 52.6 (#288), Mistral 7B: 54.2 (#280)

Instruction Following benchmarks
BenchmarkMistralMistral 7B
LMArena Instruction Following11521060
IFEval56.8%—

Long Context Mistral leads

Mistral: 35.0 (#245), Mistral 7B: 32.2 (#271)

Long Context benchmarks
BenchmarkMistralMistral 7B
LMArena Longer Query11531060

Writing & Preference Mistral leads

Mistral: 37.0 (#260), Mistral 7B: 30.7 (#286)

Writing & Preference benchmarks
BenchmarkMistralMistral 7B
LMArena Text11651090
LMArena Creative Writing11581068
LMArena Multi-Turn11471062
WildBench66%—

Frequently asked questions

Is Mistral better than Mistral 7B?

Mistral is the stronger model overall, scoring 29.9 to 23.0 on the Noometry Index.

Is Mistral or Mistral 7B better for coding?

Mistral scores higher on coding benchmarks: 33.8 versus 26.4 in the Noometry coding category.

How many benchmarks do Mistral and Mistral 7B share?

16 benchmarks have published results for both models. Mistral has 22 scored results on Noometry and Mistral 7B has 37.

Related comparisons

Go deeper