Model comparison

Mistral vs Mixtral 8x7B

Mistral is the stronger model overall, scoring 29.9 to 27.1 on the Noometry Index.

Last verified . 22 shared benchmarks.

Mistral Mistral AI

29.9

Rank #303 Confirmed

Mixtral 8x7B Mistral AI

27.1

Rank #334 Confirmed

Summary

  • They share 22 benchmarks with published results for both. Mistral scores higher in 8 categories and Mixtral 8x7B in 0 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Mistral leads 16.6 to 11.0.
  • The biggest single-benchmark swing is MMLU-Pro: 27.7% for Mistral and 33.5% for Mixtral 8x7B.
  • Mixtral 8x7B has downloadable open weights; the other is API-only.

Side by side

Mistral and Mixtral 8x7B specifications
MistralMixtral 8x7B
ProviderMistral AIMistral AI
Noometry Index29.927.1
Released—2023-12-11
WeightsProprietaryOpen
Context window—32K
Max output—32K
Input $ / M tokens—$0.70
Output $ / M tokens—$0.70
Results tracked2238

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mistral leads

Mistral: 33.8 (#250), Mixtral 8x7B: 32.8 (#269)

Coding benchmarks
BenchmarkMistralMixtral 8x7B
LMArena Coding11621126
HumanEval+—39.6%
MBPP+—49.7%

Reasoning Mistral leads

Mistral: 22.2 (#200), Mixtral 8x7B: 18.2 (#285)

Reasoning benchmarks
BenchmarkMistralMixtral 8x7B
LMArena Hard Prompts11491115
DTBench—49.6%
Adversarial NLI—55.2%
Epoch Capabilities Index—118.47
ForecastBench—56.3
HellaSwag—86.7%
PIQA—83.6%
WinoGrande—77.2%

Math Mistral leads

Mistral: 22.3 (#278), Mixtral 8x7B: 18.8 (#289)

Math benchmarks
BenchmarkMistralMixtral 8x7B
Omni-MATH7.2%10.5%
LMArena Math11801147
MATH Level 5—10%
GSM8K—74.4%

Knowledge Mistral leads

Mistral: 16.6 (#288), Mixtral 8x7B: 11.0 (#301)

Knowledge benchmarks
BenchmarkMistralMixtral 8x7B
MMLU-Pro27.7%33.5%
GPQA (HELM)30.3%29.6%
LMArena Expert11251088
GPQA Diamond—30.6%
ARC (AI2) Challenge—87.3%
MMLU—70.6%
OpenBookQA—85.8%
TriviaQA—82.2%

Multilingual Mistral leads

Mistral: 32.8 (#254), Mixtral 8x7B: 29.6 (#266)

Multilingual benchmarks
BenchmarkMistralMixtral 8x7B
LMArena Non-English11291077
LMArena Chinese11091055
LMArena French11801166
LMArena German11551114
LMArena Japanese1013931
LMArena Korean1032968
LMArena Russian11681090
LMArena Spanish11431111

Instruction Following Mistral leads

Mistral: 52.6 (#288), Mixtral 8x7B: 51.0 (#297)

Instruction Following benchmarks
BenchmarkMistralMixtral 8x7B
IFEval56.8%57.5%
LMArena Instruction Following11521109

Long Context Mistral leads

Mistral: 35.0 (#245), Mixtral 8x7B: 33.4 (#260)

Long Context benchmarks
BenchmarkMistralMixtral 8x7B
LMArena Longer Query11531103

Writing & Preference Mistral leads

Mistral: 37.0 (#260), Mixtral 8x7B: 34.2 (#270)

Writing & Preference benchmarks
BenchmarkMistralMixtral 8x7B
LMArena Text11651132
LMArena Creative Writing11581109
WildBench66%67.3%
LMArena Multi-Turn11471115

Frequently asked questions

Is Mistral better than Mixtral 8x7B?

Mistral is the stronger model overall, scoring 29.9 to 27.1 on the Noometry Index.

Is Mistral or Mixtral 8x7B better for coding?

Mistral scores higher on coding benchmarks: 33.8 versus 32.8 in the Noometry coding category.

How many benchmarks do Mistral and Mixtral 8x7B share?

22 benchmarks have published results for both models. Mistral has 22 scored results on Noometry and Mixtral 8x7B has 38.

Related comparisons

Go deeper