Model comparison

Mistral vs Yi-34B

Mistral is the stronger model overall, scoring 29.9 to 27.8 on the Noometry Index.

Last verified . 17 shared benchmarks.

Mistral Mistral AI

29.9

Rank #303 Confirmed

Yi-34B 01.AI

27.8

Rank #329 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Mistral scores higher in 7 categories and Yi-34B in 1 category; 7 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Mistral leads 16.6 to 7.5.
  • Yi-34B has downloadable open weights; the other is API-only.

Side by side

Mistral and Yi-34B specifications
MistralYi-34B
ProviderMistral AI01.AI
Noometry Index29.927.8
Released—2023-11-02
WeightsProprietaryOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked2223

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mistral leads

Mistral: 33.8 (#250), Yi-34B: 32.3 (#274)

Coding benchmarks
BenchmarkMistralYi-34B
LMArena Coding11621112

Reasoning Mistral leads

Mistral: 22.2 (#200), Yi-34B: 21.2 (#226)

Reasoning benchmarks
BenchmarkMistralYi-34B
LMArena Hard Prompts11491104
BIG-Bench Hard—71.7%
Epoch Capabilities Index—117.39

Math Too close to call

Mistral: 22.3 (#278), Yi-34B: 21.6 (#282)

Math benchmarks
BenchmarkMistralYi-34B
LMArena Math11801114
Omni-MATH7.2%—
MATH Level 5—5.1%
GSM8K—76%

Knowledge Mistral leads

Mistral: 16.6 (#288), Yi-34B: 7.5 (#309)

Knowledge benchmarks
BenchmarkMistralYi-34B
LMArena Expert11251061
GPQA Diamond—14.7%
MMLU-Pro27.7%—
GPQA (HELM)30.3%—
MMLU—76.3%

Multilingual Mistral leads

Mistral: 32.8 (#254), Yi-34B: 29.7 (#264)

Multilingual benchmarks
BenchmarkMistralYi-34B
LMArena Non-English11291079
LMArena Chinese11091176
LMArena French11801081
LMArena German11551042
LMArena Japanese1013993
LMArena Korean1032959
LMArena Russian11681050
LMArena Spanish11431070

Instruction Following Yi-34B leads

Mistral: 52.6 (#288), Yi-34B: 56.2 (#274)

Instruction Following benchmarks
BenchmarkMistralYi-34B
LMArena Instruction Following11521091
IFEval56.8%—

Long Context Mistral leads

Mistral: 35.0 (#245), Yi-34B: 33.2 (#264)

Long Context benchmarks
BenchmarkMistralYi-34B
LMArena Longer Query11531094

Writing & Preference Mistral leads

Mistral: 37.0 (#260), Yi-34B: 34.1 (#273)

Writing & Preference benchmarks
BenchmarkMistralYi-34B
LMArena Text11651129
LMArena Creative Writing11581108
LMArena Multi-Turn11471113
WildBench66%—

Frequently asked questions

Is Mistral better than Yi-34B?

Mistral is the stronger model overall, scoring 29.9 to 27.8 on the Noometry Index.

Is Mistral or Yi-34B better for coding?

Mistral scores higher on coding benchmarks: 33.8 versus 32.3 in the Noometry coding category.

How many benchmarks do Mistral and Yi-34B share?

17 benchmarks have published results for both models. Mistral has 22 scored results on Noometry and Yi-34B has 23.

Related comparisons

Go deeper