Model comparison

MiMo-V2-Pro vs Phi-4 Mini

MiMo-V2-Pro is the stronger model overall, scoring 43.0 to 30.9 on the Noometry Index. Phi-4 Mini costs 4.1× less per token, which makes it the better buy when MiMo-V2-Pro's lead doesn't matter for your workload.

Last verified . 0 shared benchmarks.

MiMo-V2-Pro Xiaomi

43.0

Rank #103 Confirmed

Phi-4 Mini Microsoft

30.9

Rank #283 Reported

Summary

  • The widest gap is in knowledge, where MiMo-V2-Pro leads 41.4 to 25.3.
  • Phi-4 Mini is cheaper at $0.075 / $0.30 per million input/output tokens, against $0.43 / $0.87 for MiMo-V2-Pro.
  • MiMo-V2-Pro accepts more context: 1.05M tokens versus 128K.
  • Phi-4 Mini has downloadable open weights; the other is API-only.

Side by side

MiMo-V2-Pro and Phi-4 Mini specifications
MiMo-V2-ProPhi-4 Mini
ProviderXiaomiMicrosoft
Noometry Index43.030.9
Released2026-03-182024-12-11
WeightsProprietaryOpen
Context window1.05M128K
Max output131K4K
Input $ / M tokens$0.43$0.075
Output $ / M tokens$0.87$0.30
Results tracked233

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding MiMo-V2-Pro leads

MiMo-V2-Pro: 43.8 (#83), Phi-4 Mini: 28.1 (#317)

Coding benchmarks
BenchmarkMiMo-V2-ProPhi-4 Mini
LMArena WebDev1433—
SciCode—10.8%
LMArena Coding1476—
ALE-Bench785.17—

Reasoning Too close to call

MiMo-V2-Pro: 22.1 (#206), Phi-4 Mini: 22.4 (#195)

Reasoning benchmarks
BenchmarkMiMo-V2-ProPhi-4 Mini
NYT Connections (extended)25.8%—
CritPt—0%
Thematic Generalization45.9%—
LMArena Hard Prompts1457—

Math Not comparable

MiMo-V2-Pro: 39.5 (#102), Phi-4 Mini: —

Math benchmarks
BenchmarkMiMo-V2-ProPhi-4 Mini
LMArena Math1447—

Knowledge MiMo-V2-Pro leads

MiMo-V2-Pro: 41.4 (#111), Phi-4 Mini: 25.3 (#262)

Knowledge benchmarks
BenchmarkMiMo-V2-ProPhi-4 Mini
Vectara Hallucination Rate—23.5%
LMArena Expert1478—

Multilingual Not comparable

MiMo-V2-Pro: 52.7 (#81), Phi-4 Mini: —

Multilingual benchmarks
BenchmarkMiMo-V2-ProPhi-4 Mini
LMArena Non-English1416—
LMArena Chinese1456—
LMArena French1469—
LMArena German1417—
LMArena Japanese1366—
LMArena Korean1400—
LMArena Russian1427—
LMArena Spanish1457—

Instruction Following Not comparable

MiMo-V2-Pro: 76.0 (#49), Phi-4 Mini: —

Instruction Following benchmarks
BenchmarkMiMo-V2-ProPhi-4 Mini
LMArena Instruction Following1445—

Long Context Not comparable

MiMo-V2-Pro: 41.5 (#138), Phi-4 Mini: —

Long Context benchmarks
BenchmarkMiMo-V2-ProPhi-4 Mini
CL-bench15.7%—
CL-bench Life6.9%—
LMArena Longer Query1455—

Writing & Preference Not comparable

MiMo-V2-Pro: 62.8 (#70), Phi-4 Mini: —

Writing & Preference benchmarks
BenchmarkMiMo-V2-ProPhi-4 Mini
LMArena Text1436—
LMArena Creative Writing1415—
LMArena Multi-Turn1456—

Frequently asked questions

Is MiMo-V2-Pro better than Phi-4 Mini?

MiMo-V2-Pro is the stronger model overall, scoring 43.0 to 30.9 on the Noometry Index. Phi-4 Mini costs 4.1× less per token, which makes it the better buy when MiMo-V2-Pro's lead doesn't matter for your workload.

Which is cheaper, MiMo-V2-Pro or Phi-4 Mini?

Phi-4 Mini is cheaper. It lists at $0.075 per million input tokens and $0.30 per million output tokens; MiMo-V2-Pro lists at $0.43 and $0.87.

Is MiMo-V2-Pro or Phi-4 Mini better for coding?

MiMo-V2-Pro scores higher on coding benchmarks: 43.8 versus 28.1 in the Noometry coding category.

Which has the bigger context window?

MiMo-V2-Pro does, with 1.05M tokens against 128K.

How many benchmarks do MiMo-V2-Pro and Phi-4 Mini share?

0 benchmarks have published results for both models. MiMo-V2-Pro has 23 scored results on Noometry and Phi-4 Mini has 3.

Related comparisons

Go deeper