Model comparison

Mixtral 8x22B vs Yi-Lightning

Yi-Lightning is the stronger model overall, scoring 37.1 to 27.1 on the Noometry Index.

Last verified . 17 shared benchmarks.

Mixtral 8x22B Mistral AI

27.1

Rank #333 Confirmed

Yi-Lightning 01.AI

37.1

Rank #209 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Mixtral 8x22B scores higher in 0 categories and Yi-Lightning in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Yi-Lightning leads 35.4 to 15.1.
  • Mixtral 8x22B has downloadable open weights; the other is API-only.

Side by side

Mixtral 8x22B and Yi-Lightning specifications
Mixtral 8x22BYi-Lightning
ProviderMistral AI01.AI
Noometry Index27.137.1
Released2024-04-172024-12-02
WeightsOpenProprietary
Context window64K—
Max output64K—
Input $ / M tokens$2—
Output $ / M tokens$6—
Results tracked3418

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Yi-Lightning leads

Mixtral 8x22B: 24.2 (#329), Yi-Lightning: 27.4 (#320)

Coding benchmarks
BenchmarkMixtral 8x22BYi-Lightning
LMArena Coding11661312
Aider Polyglot—12.9%
WeirdML3.2%—
BigCodeBench Instruct40.6%—
BigCodeBench Complete50.2%—
HumanEval+72%—
MBPP+64.3%—

Agentic & Tool Use Not comparable

Mixtral 8x22B: 23.1 (#127), Yi-Lightning: —

Agentic & Tool Use benchmarks
BenchmarkMixtral 8x22BYi-Lightning
Cybench7.5%—

Reasoning Yi-Lightning leads

Mixtral 8x22B: 19.9 (#248), Yi-Lightning: 25.9 (#139)

Reasoning benchmarks
BenchmarkMixtral 8x22BYi-Lightning
LMArena Hard Prompts11501302
DTBench55.1%—
Epoch Capabilities Index122.03—
ForecastBench56.3—

Math Yi-Lightning leads

Mixtral 8x22B: 22.9 (#275), Yi-Lightning: 36.2 (#172)

Math benchmarks
BenchmarkMixtral 8x22BYi-Lightning
LMArena Math11841300
Omni-MATH16.3%—
MATH Level 524.2%—

Knowledge Yi-Lightning leads

Mixtral 8x22B: 15.1 (#293), Yi-Lightning: 35.4 (#185)

Knowledge benchmarks
BenchmarkMixtral 8x22BYi-Lightning
LMArena Expert11131286
GPQA Diamond34.1%—
MMLU-Pro46%—
GPQA (HELM)33.4%—
MMLU77.8%—

Multilingual Yi-Lightning leads

Mixtral 8x22B: 32.8 (#255), Yi-Lightning: 42.2 (#196)

Multilingual benchmarks
BenchmarkMixtral 8x22BYi-Lightning
LMArena Non-English11281269
LMArena Chinese11161320
LMArena French11661305
LMArena German11411267
LMArena Japanese10371228
LMArena Korean10571192
LMArena Russian11581255
LMArena Spanish11511315

Instruction Following Yi-Lightning leads

Mixtral 8x22B: 57.7 (#266), Yi-Lightning: 67.4 (#196)

Instruction Following benchmarks
BenchmarkMixtral 8x22BYi-Lightning
LMArena Instruction Following11471278
IFEval72.4%—

Long Context Yi-Lightning leads

Mixtral 8x22B: 34.7 (#247), Yi-Lightning: 39.4 (#181)

Long Context benchmarks
BenchmarkMixtral 8x22BYi-Lightning
LMArena Longer Query11441297

Writing & Preference Yi-Lightning leads

Mixtral 8x22B: 36.9 (#262), Yi-Lightning: 50.2 (#184)

Writing & Preference benchmarks
BenchmarkMixtral 8x22BYi-Lightning
LMArena Text11621302
LMArena Creative Writing11411280
LMArena Multi-Turn11301311
WildBench71.1%—

Frequently asked questions

Is Mixtral 8x22B better than Yi-Lightning?

Yi-Lightning is the stronger model overall, scoring 37.1 to 27.1 on the Noometry Index.

Is Mixtral 8x22B or Yi-Lightning better for coding?

Yi-Lightning scores higher on coding benchmarks: 27.4 versus 24.2 in the Noometry coding category.

How many benchmarks do Mixtral 8x22B and Yi-Lightning share?

17 benchmarks have published results for both models. Mixtral 8x22B has 34 scored results on Noometry and Yi-Lightning has 18.

Related comparisons

Go deeper