Model comparison

MiMo-V2-Pro vs Phi 3 Mini 4k Instruct June 2024

MiMo-V2-Pro is the stronger model overall, scoring 43.0 to 31.3 on the Noometry Index.

Last verified . 15 shared benchmarks.

MiMo-V2-Pro Xiaomi

43.0

Rank #103 Confirmed

Summary

  • They share 15 benchmarks with published results for both. MiMo-V2-Pro scores higher in 8 categories and Phi 3 Mini 4k Instruct June 2024 in 0 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where MiMo-V2-Pro leads 62.8 to 29.6.
  • Phi 3 Mini 4k Instruct June 2024 has downloadable open weights; the other is API-only.

Side by side

MiMo-V2-Pro and Phi 3 Mini 4k Instruct June 2024 specifications
MiMo-V2-ProPhi 3 Mini 4k Instruct June 2024
ProviderXiaomiMicrosoft
Noometry Index43.031.3
Released2026-03-18—
WeightsProprietaryOpen
Context window1.05M—
Max output131K—
Input $ / M tokens$0.43—
Output $ / M tokens$0.87—
Results tracked2315

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding MiMo-V2-Pro leads

MiMo-V2-Pro: 43.8 (#83), Phi 3 Mini 4k Instruct June 2024: 31.8 (#279)

Coding benchmarks
BenchmarkMiMo-V2-ProPhi 3 Mini 4k Instruct June 2024
LMArena Coding14761093
LMArena WebDev1433—
ALE-Bench785.17—

Reasoning MiMo-V2-Pro leads

MiMo-V2-Pro: 22.1 (#206), Phi 3 Mini 4k Instruct June 2024: 20.8 (#231)

Reasoning benchmarks
BenchmarkMiMo-V2-ProPhi 3 Mini 4k Instruct June 2024
LMArena Hard Prompts14571087
NYT Connections (extended)25.8%—
Thematic Generalization45.9%—

Math MiMo-V2-Pro leads

MiMo-V2-Pro: 39.5 (#102), Phi 3 Mini 4k Instruct June 2024: 33.0 (#208)

Math benchmarks
BenchmarkMiMo-V2-ProPhi 3 Mini 4k Instruct June 2024
LMArena Math14471152

Knowledge MiMo-V2-Pro leads

MiMo-V2-Pro: 41.4 (#111), Phi 3 Mini 4k Instruct June 2024: 28.6 (#244)

Knowledge benchmarks
BenchmarkMiMo-V2-ProPhi 3 Mini 4k Instruct June 2024
LMArena Expert14781051

Multilingual MiMo-V2-Pro leads

MiMo-V2-Pro: 52.7 (#81), Phi 3 Mini 4k Instruct June 2024: 25.9 (#282)

Multilingual benchmarks
BenchmarkMiMo-V2-ProPhi 3 Mini 4k Instruct June 2024
LMArena Non-English14161013
LMArena Chinese14561033
LMArena German14171031
LMArena Japanese1366954
LMArena Korean1400880
LMArena Russian14271019
LMArena French1469—
LMArena Spanish1457—

Instruction Following MiMo-V2-Pro leads

MiMo-V2-Pro: 76.0 (#49), Phi 3 Mini 4k Instruct June 2024: 54.1 (#282)

Instruction Following benchmarks
BenchmarkMiMo-V2-ProPhi 3 Mini 4k Instruct June 2024
LMArena Instruction Following14451058

Long Context MiMo-V2-Pro leads

MiMo-V2-Pro: 41.5 (#138), Phi 3 Mini 4k Instruct June 2024: 31.7 (#277)

Long Context benchmarks
BenchmarkMiMo-V2-ProPhi 3 Mini 4k Instruct June 2024
LMArena Longer Query14551042
CL-bench15.7%—
CL-bench Life6.9%—

Writing & Preference MiMo-V2-Pro leads

MiMo-V2-Pro: 62.8 (#70), Phi 3 Mini 4k Instruct June 2024: 29.6 (#294)

Writing & Preference benchmarks
BenchmarkMiMo-V2-ProPhi 3 Mini 4k Instruct June 2024
LMArena Text14361080
LMArena Creative Writing14151045
LMArena Multi-Turn14561049

Frequently asked questions

Is MiMo-V2-Pro better than Phi 3 Mini 4k Instruct June 2024?

MiMo-V2-Pro is the stronger model overall, scoring 43.0 to 31.3 on the Noometry Index.

Is MiMo-V2-Pro or Phi 3 Mini 4k Instruct June 2024 better for coding?

MiMo-V2-Pro scores higher on coding benchmarks: 43.8 versus 31.8 in the Noometry coding category.

How many benchmarks do MiMo-V2-Pro and Phi 3 Mini 4k Instruct June 2024 share?

15 benchmarks have published results for both models. MiMo-V2-Pro has 23 scored results on Noometry and Phi 3 Mini 4k Instruct June 2024 has 15.

Related comparisons

Go deeper