Model comparison

DeepSeek Coder 33B vs Mixtral 8x7B

Mixtral 8x7B has enough public results to be ranked (#334); DeepSeek Coder 33B does not yet, so treat this comparison as directional.

Last verified . 7 shared benchmarks.

DeepSeek Coder 33B DeepSeek

38.9

Unranked Sparse

Mixtral 8x7B Mistral AI

27.1

Rank #334 Confirmed

Summary

  • They share 7 benchmarks with published results for both. DeepSeek Coder 33B scores higher in 1 category and Mixtral 8x7B in 0 categories; one gap is clear of the uncertainty.
  • The widest gap is in coding, where DeepSeek Coder 33B leads 38.0 to 32.8.

Side by side

DeepSeek Coder 33B and Mixtral 8x7B specifications
DeepSeek Coder 33BMixtral 8x7B
ProviderDeepSeekMistral AI
Noometry Index38.927.1
Released2023-11-022023-12-11
WeightsOpenOpen
Context window—32K
Max output—32K
Input $ / M tokens—$0.70
Output $ / M tokens—$0.70
Results tracked938

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding DeepSeek Coder 33B leads

DeepSeek Coder 33B: 38.0 (#184), Mixtral 8x7B: 32.8 (#269)

Coding benchmarks
BenchmarkDeepSeek Coder 33BMixtral 8x7B
HumanEval+75%39.6%
MBPP+70.1%49.7%
BigCodeBench Instruct42%—
LMArena Coding—1126
BigCodeBench Complete51.1%—

Reasoning Not comparable

DeepSeek Coder 33B: —, Mixtral 8x7B: 18.2 (#285)

Reasoning benchmarks
BenchmarkDeepSeek Coder 33BMixtral 8x7B
Epoch Capabilities Index96.32118.47
WinoGrande62%77.2%
LMArena Hard Prompts—1115
DTBench—49.6%
Adversarial NLI—55.2%
ForecastBench—56.3
HellaSwag—86.7%
PIQA—83.6%

Math Not comparable

DeepSeek Coder 33B: —, Mixtral 8x7B: 18.8 (#289)

Math benchmarks
BenchmarkDeepSeek Coder 33BMixtral 8x7B
GSM8K35.4%74.4%
Omni-MATH—10.5%
LMArena Math—1147
MATH Level 5—10%

Knowledge Not comparable

DeepSeek Coder 33B: —, Mixtral 8x7B: 11.0 (#301)

Knowledge benchmarks
BenchmarkDeepSeek Coder 33BMixtral 8x7B
ARC (AI2) Challenge42.2%87.3%
MMLU39.4%70.6%
GPQA Diamond—30.6%
MMLU-Pro—33.5%
GPQA (HELM)—29.6%
LMArena Expert—1088
OpenBookQA—85.8%
TriviaQA—82.2%

Multilingual Not comparable

DeepSeek Coder 33B: —, Mixtral 8x7B: 29.6 (#266)

Multilingual benchmarks
BenchmarkDeepSeek Coder 33BMixtral 8x7B
LMArena Non-English—1077
LMArena Chinese—1055
LMArena French—1166
LMArena German—1114
LMArena Japanese—931
LMArena Korean—968
LMArena Russian—1090
LMArena Spanish—1111

Instruction Following Not comparable

DeepSeek Coder 33B: —, Mixtral 8x7B: 51.0 (#297)

Instruction Following benchmarks
BenchmarkDeepSeek Coder 33BMixtral 8x7B
IFEval—57.5%
LMArena Instruction Following—1109

Long Context Not comparable

DeepSeek Coder 33B: —, Mixtral 8x7B: 33.4 (#260)

Long Context benchmarks
BenchmarkDeepSeek Coder 33BMixtral 8x7B
LMArena Longer Query—1103

Writing & Preference Not comparable

DeepSeek Coder 33B: —, Mixtral 8x7B: 34.2 (#270)

Writing & Preference benchmarks
BenchmarkDeepSeek Coder 33BMixtral 8x7B
LMArena Text—1132
LMArena Creative Writing—1109
WildBench—67.3%
LMArena Multi-Turn—1115

Frequently asked questions

Is DeepSeek Coder 33B better than Mixtral 8x7B?

Mixtral 8x7B has enough public results to be ranked (#334); DeepSeek Coder 33B does not yet, so treat this comparison as directional.

Is DeepSeek Coder 33B or Mixtral 8x7B better for coding?

DeepSeek Coder 33B scores higher on coding benchmarks: 38.0 versus 32.8 in the Noometry coding category.

How many benchmarks do DeepSeek Coder 33B and Mixtral 8x7B share?

7 benchmarks have published results for both models. DeepSeek Coder 33B has 9 scored results on Noometry and Mixtral 8x7B has 38.

Related comparisons

Go deeper