Model comparison

Mistral Medium 3.5 vs Mixtral 8x7B

Mistral Medium 3.5 is the stronger model overall, scoring 40.2 to 27.1 on the Noometry Index. Mixtral 8x7B costs 4.3× less per token, which makes it the better buy when Mistral Medium 3.5's lead doesn't matter for your workload.

Last verified . 17 shared benchmarks.

Mistral Medium 3.5 Mistral AI

40.2

Rank #152 Confirmed

Mixtral 8x7B Mistral AI

27.1

Rank #334 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Mistral Medium 3.5 scores higher in 7 categories and Mixtral 8x7B in 1 category; 7 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Mistral Medium 3.5 leads 40.0 to 11.0.
  • Mixtral 8x7B is cheaper at $0.70 / $0.70 per million input/output tokens, against $1.50 / $7.50 for Mistral Medium 3.5.
  • Mistral Medium 3.5 accepts more context: 262K tokens versus 32K.

Side by side

Mistral Medium 3.5 and Mixtral 8x7B specifications
Mistral Medium 3.5Mixtral 8x7B
ProviderMistral AIMistral AI
Noometry Index40.227.1
Released—2023-12-11
WeightsOpenOpen
Context window262K32K
Max output210K32K
Input $ / M tokens$1.50$0.70
Output $ / M tokens$7.50$0.70
Results tracked2238

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mistral Medium 3.5 leads

Mistral Medium 3.5: 36.0 (#213), Mixtral 8x7B: 32.8 (#269)

Coding benchmarks
BenchmarkMistral Medium 3.5Mixtral 8x7B
LMArena Coding14611126
LMArena WebDev1264—
HumanEval+—39.6%
MBPP+—49.7%

Reasoning Too close to call

Mistral Medium 3.5: 17.3 (#295), Mixtral 8x7B: 18.2 (#285)

Reasoning benchmarks
BenchmarkMistral Medium 3.5Mixtral 8x7B
LMArena Hard Prompts14361115
Epoch Capabilities Index141.35118.47
Kagi LLM Benchmark41.4%—
NYT Connections (extended)12.9%—
DTBench—49.6%
Adversarial NLI—55.2%
ForecastBench—56.3
HellaSwag—86.7%
PIQA—83.6%
WinoGrande—77.2%

Math Mistral Medium 3.5 leads

Mistral Medium 3.5: 39.1 (#113), Mixtral 8x7B: 18.8 (#289)

Math benchmarks
BenchmarkMistral Medium 3.5Mixtral 8x7B
LMArena Math14311147
Omni-MATH—10.5%
MATH Level 5—10%
GSM8K—74.4%

Knowledge Mistral Medium 3.5 leads

Mistral Medium 3.5: 40.0 (#126), Mixtral 8x7B: 11.0 (#301)

Knowledge benchmarks
BenchmarkMistral Medium 3.5Mixtral 8x7B
LMArena Expert14321088
GPQA Diamond—30.6%
MMLU-Pro—33.5%
GPQA (HELM)—29.6%
ARC (AI2) Challenge—87.3%
MMLU—70.6%
OpenBookQA—85.8%
TriviaQA—82.2%

Multimodal Not comparable

Mistral Medium 3.5: 38.3 (#65), Mixtral 8x7B: —

Multimodal benchmarks
BenchmarkMistral Medium 3.5Mixtral 8x7B
LMArena Vision1223—

Multilingual Mistral Medium 3.5 leads

Mistral Medium 3.5: 51.9 (#100), Mixtral 8x7B: 29.6 (#266)

Multilingual benchmarks
BenchmarkMistral Medium 3.5Mixtral 8x7B
LMArena Non-English14041077
LMArena Chinese14421055
LMArena French14481166
LMArena German14511114
LMArena Korean1385968
LMArena Russian13951090
LMArena Spanish14091111
LMArena Japanese—931

Instruction Following Mistral Medium 3.5 leads

Mistral Medium 3.5: 74.6 (#90), Mixtral 8x7B: 51.0 (#297)

Instruction Following benchmarks
BenchmarkMistral Medium 3.5Mixtral 8x7B
LMArena Instruction Following14151109
IFEval—57.5%

Long Context Mistral Medium 3.5 leads

Mistral Medium 3.5: 43.2 (#103), Mixtral 8x7B: 33.4 (#260)

Long Context benchmarks
BenchmarkMistral Medium 3.5Mixtral 8x7B
LMArena Longer Query14151103

Writing & Preference Mistral Medium 3.5 leads

Mistral Medium 3.5: 58.5 (#117), Mixtral 8x7B: 34.2 (#270)

Writing & Preference benchmarks
BenchmarkMistral Medium 3.5Mixtral 8x7B
LMArena Text14211132
LMArena Creative Writing13741109
LMArena Multi-Turn14231115
WildBench—67.3%
EQ-Bench 4993—

Frequently asked questions

Is Mistral Medium 3.5 better than Mixtral 8x7B?

Mistral Medium 3.5 is the stronger model overall, scoring 40.2 to 27.1 on the Noometry Index. Mixtral 8x7B costs 4.3× less per token, which makes it the better buy when Mistral Medium 3.5's lead doesn't matter for your workload.

Which is cheaper, Mistral Medium 3.5 or Mixtral 8x7B?

Mixtral 8x7B is cheaper. It lists at $0.70 per million input tokens and $0.70 per million output tokens; Mistral Medium 3.5 lists at $1.50 and $7.50.

Is Mistral Medium 3.5 or Mixtral 8x7B better for coding?

Mistral Medium 3.5 scores higher on coding benchmarks: 36.0 versus 32.8 in the Noometry coding category.

Which has the bigger context window?

Mistral Medium 3.5 does, with 262K tokens against 32K.

How many benchmarks do Mistral Medium 3.5 and Mixtral 8x7B share?

17 benchmarks have published results for both models. Mistral Medium 3.5 has 22 scored results on Noometry and Mixtral 8x7B has 38.

Related comparisons

Go deeper