Model comparison

MiMo-V2-Pro vs Step 3.7 Flash

MiMo-V2-Pro is the stronger model overall, scoring 43.0 to 37.3 on the Noometry Index.

Last verified . 2 shared benchmarks.

MiMo-V2-Pro Xiaomi

43.0

Rank #103 Confirmed

Step 3.7 Flash StepFun

37.3

Rank #207 Reported

Summary

  • They share 2 benchmarks with published results for both. MiMo-V2-Pro scores higher in 2 categories and Step 3.7 Flash in 1 category; 2 gaps are clear of the uncertainty.
  • The widest gap is in coding, where MiMo-V2-Pro leads 43.8 to 40.0.
  • The biggest single-benchmark swing is NYT Connections (extended): 25.8% for MiMo-V2-Pro and 39.7% for Step 3.7 Flash.
  • Step 3.7 Flash is cheaper at $0.18 / $1.11 per million input/output tokens, against $0.43 / $0.87 for MiMo-V2-Pro.
  • MiMo-V2-Pro accepts more context: 1.05M tokens versus 256K.
  • Step 3.7 Flash has downloadable open weights; the other is API-only.

Side by side

MiMo-V2-Pro and Step 3.7 Flash specifications
MiMo-V2-ProStep 3.7 Flash
ProviderXiaomiStepFun
Noometry Index43.037.3
Released2026-03-182026-05-29
WeightsProprietaryOpen
Context window1.05M256K
Max output131K256K
Input $ / M tokens$0.43$0.18
Output $ / M tokens$0.87$1.11
Results tracked235

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding MiMo-V2-Pro leads

MiMo-V2-Pro: 43.8 (#83), Step 3.7 Flash: 40.0 (#150)

Coding benchmarks
BenchmarkMiMo-V2-ProStep 3.7 Flash
ALE-Bench785.17694.12
LMArena WebDev1433—
SciCode—40%
LMArena Coding1476—

Reasoning Too close to call

MiMo-V2-Pro: 22.1 (#206), Step 3.7 Flash: 21.6 (#219)

Reasoning benchmarks
BenchmarkMiMo-V2-ProStep 3.7 Flash
NYT Connections (extended)25.8%39.7%
CritPt—2.3%
Thematic Generalization45.9%—
LMArena Hard Prompts1457—

Math Step 3.7 Flash leads

MiMo-V2-Pro: 39.5 (#102), Step 3.7 Flash: 42.9 (#82)

Math benchmarks
BenchmarkMiMo-V2-ProStep 3.7 Flash
MathArena Final-Answer Competitions—68.5%
LMArena Math1447—

Knowledge Not comparable

MiMo-V2-Pro: 41.4 (#111), Step 3.7 Flash: —

Knowledge benchmarks
BenchmarkMiMo-V2-ProStep 3.7 Flash
LMArena Expert1478—

Multilingual Not comparable

MiMo-V2-Pro: 52.7 (#81), Step 3.7 Flash: —

Multilingual benchmarks
BenchmarkMiMo-V2-ProStep 3.7 Flash
LMArena Non-English1416—
LMArena Chinese1456—
LMArena French1469—
LMArena German1417—
LMArena Japanese1366—
LMArena Korean1400—
LMArena Russian1427—
LMArena Spanish1457—

Instruction Following Not comparable

MiMo-V2-Pro: 76.0 (#49), Step 3.7 Flash: —

Instruction Following benchmarks
BenchmarkMiMo-V2-ProStep 3.7 Flash
LMArena Instruction Following1445—

Long Context Not comparable

MiMo-V2-Pro: 41.5 (#138), Step 3.7 Flash: —

Long Context benchmarks
BenchmarkMiMo-V2-ProStep 3.7 Flash
CL-bench15.7%—
CL-bench Life6.9%—
LMArena Longer Query1455—

Writing & Preference Not comparable

MiMo-V2-Pro: 62.8 (#70), Step 3.7 Flash: —

Writing & Preference benchmarks
BenchmarkMiMo-V2-ProStep 3.7 Flash
LMArena Text1436—
LMArena Creative Writing1415—
LMArena Multi-Turn1456—

Frequently asked questions

Is MiMo-V2-Pro better than Step 3.7 Flash?

MiMo-V2-Pro is the stronger model overall, scoring 43.0 to 37.3 on the Noometry Index.

Which is cheaper, MiMo-V2-Pro or Step 3.7 Flash?

Step 3.7 Flash is cheaper. It lists at $0.18 per million input tokens and $1.11 per million output tokens; MiMo-V2-Pro lists at $0.43 and $0.87.

Is MiMo-V2-Pro or Step 3.7 Flash better for coding?

MiMo-V2-Pro scores higher on coding benchmarks: 43.8 versus 40.0 in the Noometry coding category.

Which has the bigger context window?

MiMo-V2-Pro does, with 1.05M tokens against 256K.

How many benchmarks do MiMo-V2-Pro and Step 3.7 Flash share?

2 benchmarks have published results for both models. MiMo-V2-Pro has 23 scored results on Noometry and Step 3.7 Flash has 5.

Related comparisons

Go deeper