Model comparison

Magistral Medium vs Qwen3.5 Max Preview

Qwen3.5 Max Preview is the stronger model overall, scoring 45.3 to 35.2 on the Noometry Index.

Last verified . 17 shared benchmarks.

Magistral Medium Mistral AI

35.2

Rank #227 Confirmed

Qwen3.5 Max Preview Alibaba (Qwen)

45.3

Rank #71 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Magistral Medium scores higher in 0 categories and Qwen3.5 Max Preview in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Qwen3.5 Max Preview leads 30.8 to 8.6.
  • Magistral Medium has downloadable open weights; the other is API-only.

Side by side

Magistral Medium and Qwen3.5 Max Preview specifications
Magistral MediumQwen3.5 Max Preview
ProviderMistral AIAlibaba (Qwen)
Noometry Index35.245.3
Released2025-03-17—
WeightsOpenProprietary
Context window262K—
Max output16K—
Input $ / M tokens$2—
Output $ / M tokens$5—
Results tracked2217

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen3.5 Max Preview leads

Magistral Medium: 39.1 (#161), Qwen3.5 Max Preview: 44.0 (#77)

Coding benchmarks
BenchmarkMagistral MediumQwen3.5 Max Preview
LMArena Coding13191487
SciCode39.2%—

Reasoning Qwen3.5 Max Preview leads

Magistral Medium: 8.6 (#348), Qwen3.5 Max Preview: 30.8 (#84)

Reasoning benchmarks
BenchmarkMagistral MediumQwen3.5 Max Preview
LMArena Hard Prompts12671483
ARC-AGI-20%—
Kagi LLM Benchmark16.2%—
ARC-AGI-16.1%—
CritPt0.3%—

Math Qwen3.5 Max Preview leads

Magistral Medium: 35.1 (#189), Qwen3.5 Max Preview: 40.1 (#94)

Math benchmarks
BenchmarkMagistral MediumQwen3.5 Max Preview
LMArena Math12501474

Knowledge Qwen3.5 Max Preview leads

Magistral Medium: 33.5 (#202), Qwen3.5 Max Preview: 41.8 (#107)

Knowledge benchmarks
BenchmarkMagistral MediumQwen3.5 Max Preview
LMArena Expert12231489

Multilingual Qwen3.5 Max Preview leads

Magistral Medium: 39.6 (#224), Qwen3.5 Max Preview: 56.2 (#22)

Multilingual benchmarks
BenchmarkMagistral MediumQwen3.5 Max Preview
LMArena Non-English12321465
LMArena Chinese12271534
LMArena French12671484
LMArena German12481487
LMArena Japanese11751495
LMArena Korean11251438
LMArena Russian12241471
LMArena Spanish12711470

Instruction Following Qwen3.5 Max Preview leads

Magistral Medium: 66.0 (#211), Qwen3.5 Max Preview: 77.0 (#31)

Instruction Following benchmarks
BenchmarkMagistral MediumQwen3.5 Max Preview
LMArena Instruction Following12541467

Long Context Qwen3.5 Max Preview leads

Magistral Medium: 39.3 (#183), Qwen3.5 Max Preview: 45.2 (#45)

Long Context benchmarks
BenchmarkMagistral MediumQwen3.5 Max Preview
LMArena Longer Query12951476

Writing & Preference Qwen3.5 Max Preview leads

Magistral Medium: 46.3 (#219), Qwen3.5 Max Preview: 66.0 (#41)

Writing & Preference benchmarks
BenchmarkMagistral MediumQwen3.5 Max Preview
LMArena Text12551470
LMArena Creative Writing12451464
LMArena Multi-Turn12751478

Frequently asked questions

Is Magistral Medium better than Qwen3.5 Max Preview?

Qwen3.5 Max Preview is the stronger model overall, scoring 45.3 to 35.2 on the Noometry Index.

Is Magistral Medium or Qwen3.5 Max Preview better for coding?

Qwen3.5 Max Preview scores higher on coding benchmarks: 44.0 versus 39.1 in the Noometry coding category.

How many benchmarks do Magistral Medium and Qwen3.5 Max Preview share?

17 benchmarks have published results for both models. Magistral Medium has 22 scored results on Noometry and Qwen3.5 Max Preview has 17.

Related comparisons

Go deeper