Model comparison

Mixtral 8x7B vs Yi-34B

Mixtral 8x7B and Yi-34B score almost the same on the Noometry Index (27.1 vs 27.8), so choose on price, context window or the category you care about most.

Last verified . 22 shared benchmarks.

Mixtral 8x7B Mistral AI

27.1

Rank #334 Confirmed

Yi-34B 01.AI

27.8

Rank #329 Confirmed

Summary

  • They share 22 benchmarks with published results for both. Mixtral 8x7B scores higher in 4 categories and Yi-34B in 4 categories; 4 gaps are clear of the uncertainty.
  • The widest gap is in instruction following, where Yi-34B leads 56.2 to 51.0.
  • The biggest single-benchmark swing is GPQA Diamond: 30.6% for Mixtral 8x7B and 14.7% for Yi-34B.

Side by side

Mixtral 8x7B and Yi-34B specifications
Mixtral 8x7BYi-34B
ProviderMistral AI01.AI
Noometry Index27.127.8
Released2023-12-112023-11-02
WeightsOpenOpen
Context window32K—
Max output32K—
Input $ / M tokens$0.70—
Output $ / M tokens$0.70—
Results tracked3823

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Mixtral 8x7B: 32.8 (#269), Yi-34B: 32.3 (#274)

Coding benchmarks
BenchmarkMixtral 8x7BYi-34B
LMArena Coding11261112
HumanEval+39.6%—
MBPP+49.7%—

Reasoning Yi-34B leads

Mixtral 8x7B: 18.2 (#285), Yi-34B: 21.2 (#226)

Reasoning benchmarks
BenchmarkMixtral 8x7BYi-34B
LMArena Hard Prompts11151104
Epoch Capabilities Index118.47117.39
DTBench49.6%—
Adversarial NLI55.2%—
BIG-Bench Hard—71.7%
ForecastBench56.3—
HellaSwag86.7%—
PIQA83.6%—
WinoGrande77.2%—

Math Yi-34B leads

Mixtral 8x7B: 18.8 (#289), Yi-34B: 21.6 (#282)

Math benchmarks
BenchmarkMixtral 8x7BYi-34B
LMArena Math11471114
MATH Level 510%5.1%
GSM8K74.4%76%
Omni-MATH10.5%—

Knowledge Mixtral 8x7B leads

Mixtral 8x7B: 11.0 (#301), Yi-34B: 7.5 (#309)

Knowledge benchmarks
BenchmarkMixtral 8x7BYi-34B
GPQA Diamond30.6%14.7%
LMArena Expert10881061
MMLU70.6%76.3%
MMLU-Pro33.5%—
GPQA (HELM)29.6%—
ARC (AI2) Challenge87.3%—
OpenBookQA85.8%—
TriviaQA82.2%—

Multilingual Too close to call

Mixtral 8x7B: 29.6 (#266), Yi-34B: 29.7 (#264)

Multilingual benchmarks
BenchmarkMixtral 8x7BYi-34B
LMArena Non-English10771079
LMArena Chinese10551176
LMArena French11661081
LMArena German11141042
LMArena Japanese931993
LMArena Korean968959
LMArena Russian10901050
LMArena Spanish11111070

Instruction Following Yi-34B leads

Mixtral 8x7B: 51.0 (#297), Yi-34B: 56.2 (#274)

Instruction Following benchmarks
BenchmarkMixtral 8x7BYi-34B
LMArena Instruction Following11091091
IFEval57.5%—

Long Context Too close to call

Mixtral 8x7B: 33.4 (#260), Yi-34B: 33.2 (#264)

Long Context benchmarks
BenchmarkMixtral 8x7BYi-34B
LMArena Longer Query11031094

Writing & Preference Too close to call

Mixtral 8x7B: 34.2 (#270), Yi-34B: 34.1 (#273)

Writing & Preference benchmarks
BenchmarkMixtral 8x7BYi-34B
LMArena Text11321129
LMArena Creative Writing11091108
LMArena Multi-Turn11151113
WildBench67.3%—

Frequently asked questions

Is Mixtral 8x7B better than Yi-34B?

Mixtral 8x7B and Yi-34B score almost the same on the Noometry Index (27.1 vs 27.8), so choose on price, context window or the category you care about most.

Is Mixtral 8x7B or Yi-34B better for coding?

They score almost the same on coding (32.8 vs 32.3); test both on your own repository before choosing.

How many benchmarks do Mixtral 8x7B and Yi-34B share?

22 benchmarks have published results for both models. Mixtral 8x7B has 38 scored results on Noometry and Yi-34B has 23.

Related comparisons

Go deeper