Model comparison

Magistral Medium vs MiMo-V2-Flash

MiMo-V2-Flash is the stronger model overall, scoring 41.3 to 35.2 on the Noometry Index.

Last verified . 19 shared benchmarks.

Magistral Medium Mistral AI

35.2

Rank #227 Confirmed

MiMo-V2-Flash Xiaomi

41.3

Rank #138 Confirmed

Summary

  • They share 19 benchmarks with published results for both. Magistral Medium scores higher in 1 category and MiMo-V2-Flash in 7 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where MiMo-V2-Flash leads 24.9 to 8.6.
  • The biggest single-benchmark swing is SciCode: 39.2% for Magistral Medium and 25.9% for MiMo-V2-Flash.
  • MiMo-V2-Flash is cheaper at $0.14 / $0.28 per million input/output tokens, against $2 / $5 for Magistral Medium.

Side by side

Magistral Medium and MiMo-V2-Flash specifications
Magistral MediumMiMo-V2-Flash
ProviderMistral AIXiaomi
Noometry Index35.241.3
Released2025-03-172025-12-16
WeightsOpenOpen
Context window262K262K
Max output16K66K
Input $ / M tokens$2$0.14
Output $ / M tokens$5$0.28
Results tracked2221

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Magistral Medium leads

Magistral Medium: 39.1 (#161), MiMo-V2-Flash: 36.1 (#211)

Coding benchmarks
BenchmarkMagistral MediumMiMo-V2-Flash
SciCode39.2%25.9%
LMArena Coding13191443
LMArena WebDev—1330
ALE-Bench—737.95

Reasoning MiMo-V2-Flash leads

Magistral Medium: 8.6 (#348), MiMo-V2-Flash: 24.9 (#157)

Reasoning benchmarks
BenchmarkMagistral MediumMiMo-V2-Flash
CritPt0.3%0%
LMArena Hard Prompts12671420
ARC-AGI-20%—
Kagi LLM Benchmark16.2%—
ARC-AGI-16.1%—

Math MiMo-V2-Flash leads

Magistral Medium: 35.1 (#189), MiMo-V2-Flash: 38.3 (#139)

Math benchmarks
BenchmarkMagistral MediumMiMo-V2-Flash
LMArena Math12501396

Knowledge MiMo-V2-Flash leads

Magistral Medium: 33.5 (#202), MiMo-V2-Flash: 39.7 (#131)

Knowledge benchmarks
BenchmarkMagistral MediumMiMo-V2-Flash
LMArena Expert12231425

Multilingual MiMo-V2-Flash leads

Magistral Medium: 39.6 (#224), MiMo-V2-Flash: 51.0 (#113)

Multilingual benchmarks
BenchmarkMagistral MediumMiMo-V2-Flash
LMArena Non-English12321392
LMArena Chinese12271462
LMArena French12671429
LMArena German12481395
LMArena Japanese11751325
LMArena Korean11251358
LMArena Russian12241387
LMArena Spanish12711420

Instruction Following MiMo-V2-Flash leads

Magistral Medium: 66.0 (#211), MiMo-V2-Flash: 73.5 (#120)

Instruction Following benchmarks
BenchmarkMagistral MediumMiMo-V2-Flash
LMArena Instruction Following12541392

Long Context MiMo-V2-Flash leads

Magistral Medium: 39.3 (#183), MiMo-V2-Flash: 43.0 (#110)

Long Context benchmarks
BenchmarkMagistral MediumMiMo-V2-Flash
LMArena Longer Query12951409

Writing & Preference MiMo-V2-Flash leads

Magistral Medium: 46.3 (#219), MiMo-V2-Flash: 59.7 (#106)

Writing & Preference benchmarks
BenchmarkMagistral MediumMiMo-V2-Flash
LMArena Text12551411
LMArena Creative Writing12451375
LMArena Multi-Turn12751404

Frequently asked questions

Is Magistral Medium better than MiMo-V2-Flash?

MiMo-V2-Flash is the stronger model overall, scoring 41.3 to 35.2 on the Noometry Index.

Which is cheaper, Magistral Medium or MiMo-V2-Flash?

MiMo-V2-Flash is cheaper. It lists at $0.14 per million input tokens and $0.28 per million output tokens; Magistral Medium lists at $2 and $5.

Is Magistral Medium or MiMo-V2-Flash better for coding?

Magistral Medium scores higher on coding benchmarks: 39.1 versus 36.1 in the Noometry coding category.

Which has the bigger context window?

Both accept 262K tokens.

How many benchmarks do Magistral Medium and MiMo-V2-Flash share?

19 benchmarks have published results for both models. Magistral Medium has 22 scored results on Noometry and MiMo-V2-Flash has 21.

Related comparisons

Go deeper