Model comparison

Mixtral 8x7B vs Qwen3.5 Max Preview

Qwen3.5 Max Preview is the stronger model overall, scoring 45.3 to 27.1 on the Noometry Index.

Last verified . 17 shared benchmarks.

Mixtral 8x7B Mistral AI

27.1

Rank #334 Confirmed

Qwen3.5 Max Preview Alibaba (Qwen)

45.3

Rank #71 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Mixtral 8x7B scores higher in 0 categories and Qwen3.5 Max Preview in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Qwen3.5 Max Preview leads 66.0 to 34.2.
  • Mixtral 8x7B has downloadable open weights; the other is API-only.

Side by side

Mixtral 8x7B and Qwen3.5 Max Preview specifications
Mixtral 8x7BQwen3.5 Max Preview
ProviderMistral AIAlibaba (Qwen)
Noometry Index27.145.3
Released2023-12-11—
WeightsOpenProprietary
Context window32K—
Max output32K—
Input $ / M tokens$0.70—
Output $ / M tokens$0.70—
Results tracked3817

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen3.5 Max Preview leads

Mixtral 8x7B: 32.8 (#269), Qwen3.5 Max Preview: 44.0 (#77)

Coding benchmarks
BenchmarkMixtral 8x7BQwen3.5 Max Preview
LMArena Coding11261487
HumanEval+39.6%—
MBPP+49.7%—

Reasoning Qwen3.5 Max Preview leads

Mixtral 8x7B: 18.2 (#285), Qwen3.5 Max Preview: 30.8 (#84)

Reasoning benchmarks
BenchmarkMixtral 8x7BQwen3.5 Max Preview
LMArena Hard Prompts11151483
DTBench49.6%—
Adversarial NLI55.2%—
Epoch Capabilities Index118.47—
ForecastBench56.3—
HellaSwag86.7%—
PIQA83.6%—
WinoGrande77.2%—

Math Qwen3.5 Max Preview leads

Mixtral 8x7B: 18.8 (#289), Qwen3.5 Max Preview: 40.1 (#94)

Math benchmarks
BenchmarkMixtral 8x7BQwen3.5 Max Preview
LMArena Math11471474
Omni-MATH10.5%—
MATH Level 510%—
GSM8K74.4%—

Knowledge Qwen3.5 Max Preview leads

Mixtral 8x7B: 11.0 (#301), Qwen3.5 Max Preview: 41.8 (#107)

Knowledge benchmarks
BenchmarkMixtral 8x7BQwen3.5 Max Preview
LMArena Expert10881489
GPQA Diamond30.6%—
MMLU-Pro33.5%—
GPQA (HELM)29.6%—
ARC (AI2) Challenge87.3%—
MMLU70.6%—
OpenBookQA85.8%—
TriviaQA82.2%—

Multilingual Qwen3.5 Max Preview leads

Mixtral 8x7B: 29.6 (#266), Qwen3.5 Max Preview: 56.2 (#22)

Multilingual benchmarks
BenchmarkMixtral 8x7BQwen3.5 Max Preview
LMArena Non-English10771465
LMArena Chinese10551534
LMArena French11661484
LMArena German11141487
LMArena Japanese9311495
LMArena Korean9681438
LMArena Russian10901471
LMArena Spanish11111470

Instruction Following Qwen3.5 Max Preview leads

Mixtral 8x7B: 51.0 (#297), Qwen3.5 Max Preview: 77.0 (#31)

Instruction Following benchmarks
BenchmarkMixtral 8x7BQwen3.5 Max Preview
LMArena Instruction Following11091467
IFEval57.5%—

Long Context Qwen3.5 Max Preview leads

Mixtral 8x7B: 33.4 (#260), Qwen3.5 Max Preview: 45.2 (#45)

Long Context benchmarks
BenchmarkMixtral 8x7BQwen3.5 Max Preview
LMArena Longer Query11031476

Writing & Preference Qwen3.5 Max Preview leads

Mixtral 8x7B: 34.2 (#270), Qwen3.5 Max Preview: 66.0 (#41)

Writing & Preference benchmarks
BenchmarkMixtral 8x7BQwen3.5 Max Preview
LMArena Text11321470
LMArena Creative Writing11091464
LMArena Multi-Turn11151478
WildBench67.3%—

Frequently asked questions

Is Mixtral 8x7B better than Qwen3.5 Max Preview?

Qwen3.5 Max Preview is the stronger model overall, scoring 45.3 to 27.1 on the Noometry Index.

Is Mixtral 8x7B or Qwen3.5 Max Preview better for coding?

Qwen3.5 Max Preview scores higher on coding benchmarks: 44.0 versus 32.8 in the Noometry coding category.

How many benchmarks do Mixtral 8x7B and Qwen3.5 Max Preview share?

17 benchmarks have published results for both models. Mixtral 8x7B has 38 scored results on Noometry and Qwen3.5 Max Preview has 17.

Related comparisons

Go deeper