Model comparison

Ministral 8B vs Mixtral 8x7B

Ministral 8B is the stronger model overall, scoring 28.2 to 27.1 on the Noometry Index.

Last verified . 15 shared benchmarks.

Ministral 8B Mistral AI

28.2

Rank #325 Confirmed

Mixtral 8x7B Mistral AI

27.1

Rank #334 Confirmed

Summary

  • They share 15 benchmarks with published results for both. Ministral 8B scores higher in 8 categories and Mixtral 8x7B in 0 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in instruction following, where Ministral 8B leads 60.5 to 51.0.
  • Ministral 8B is cheaper at $0.15 / $0.15 per million input/output tokens, against $0.70 / $0.70 for Mixtral 8x7B.
  • Ministral 8B accepts more context: 262K tokens versus 32K.

Side by side

Ministral 8B and Mixtral 8x7B specifications
Ministral 8BMixtral 8x7B
ProviderMistral AIMistral AI
Noometry Index28.227.1
Released2024-10-012023-12-11
WeightsOpenOpen
Context window262K32K
Max output262K32K
Input $ / M tokens$0.15$0.70
Output $ / M tokens$0.15$0.70
Results tracked1738

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Ministral 8B leads

Ministral 8B: 35.0 (#230), Mixtral 8x7B: 32.8 (#269)

Coding benchmarks
BenchmarkMinistral 8BMixtral 8x7B
LMArena Coding12021126
HumanEval+—39.6%
MBPP+—49.7%

Agentic & Tool Use Not comparable

Ministral 8B: 16.4 (#148), Mixtral 8x7B: —

Agentic & Tool Use benchmarks
BenchmarkMinistral 8BMixtral 8x7B
Berkeley Function Calling Leaderboard11.1%—

Reasoning Too close to call

Ministral 8B: 18.4 (#281), Mixtral 8x7B: 18.2 (#285)

Reasoning benchmarks
BenchmarkMinistral 8BMixtral 8x7B
LMArena Hard Prompts11911115
DTBench45.7%49.6%
Adversarial NLI—55.2%
Epoch Capabilities Index—118.47
ForecastBench—56.3
HellaSwag—86.7%
PIQA—83.6%
WinoGrande—77.2%

Math Ministral 8B leads

Ministral 8B: 25.7 (#267), Mixtral 8x7B: 18.8 (#289)

Math benchmarks
BenchmarkMinistral 8BMixtral 8x7B
LMArena Math11881147
MATH Level 514.9%10%
Omni-MATH—10.5%
GSM8K—74.4%

Knowledge Ministral 8B leads

Ministral 8B: 12.6 (#297), Mixtral 8x7B: 11.0 (#301)

Knowledge benchmarks
BenchmarkMinistral 8BMixtral 8x7B
GPQA Diamond27.1%30.6%
LMArena Expert11701088
MMLU-Pro—33.5%
Vectara Hallucination Rate7.4%—
GPQA (HELM)—29.6%
ARC (AI2) Challenge—87.3%
MMLU—70.6%
OpenBookQA—85.8%
TriviaQA—82.2%

Multilingual Ministral 8B leads

Ministral 8B: 35.1 (#247), Mixtral 8x7B: 29.6 (#266)

Multilingual benchmarks
BenchmarkMinistral 8BMixtral 8x7B
LMArena Non-English11651077
LMArena Chinese11931055
LMArena Russian11951090
LMArena French—1166
LMArena German—1114
LMArena Japanese—931
LMArena Korean—968
LMArena Spanish—1111

Instruction Following Ministral 8B leads

Ministral 8B: 60.5 (#250), Mixtral 8x7B: 51.0 (#297)

Instruction Following benchmarks
BenchmarkMinistral 8BMixtral 8x7B
LMArena Instruction Following11611109
IFEval—57.5%

Long Context Ministral 8B leads

Ministral 8B: 36.7 (#227), Mixtral 8x7B: 33.4 (#260)

Long Context benchmarks
BenchmarkMinistral 8BMixtral 8x7B
LMArena Longer Query12121103

Writing & Preference Ministral 8B leads

Ministral 8B: 39.6 (#246), Mixtral 8x7B: 34.2 (#270)

Writing & Preference benchmarks
BenchmarkMinistral 8BMixtral 8x7B
LMArena Text11911132
LMArena Creative Writing11751109
LMArena Multi-Turn11661115
WildBench—67.3%

Frequently asked questions

Is Ministral 8B better than Mixtral 8x7B?

Ministral 8B is the stronger model overall, scoring 28.2 to 27.1 on the Noometry Index.

Which is cheaper, Ministral 8B or Mixtral 8x7B?

Ministral 8B is cheaper. It lists at $0.15 per million input tokens and $0.15 per million output tokens; Mixtral 8x7B lists at $0.70 and $0.70.

Is Ministral 8B or Mixtral 8x7B better for coding?

Ministral 8B scores higher on coding benchmarks: 35.0 versus 32.8 in the Noometry coding category.

Which has the bigger context window?

Ministral 8B does, with 262K tokens against 32K.

How many benchmarks do Ministral 8B and Mixtral 8x7B share?

15 benchmarks have published results for both models. Ministral 8B has 17 scored results on Noometry and Mixtral 8x7B has 38.

Related comparisons

Go deeper