Model comparison

Hy3 vs MiMo-V2-Pro

Hy3 is the stronger model overall, scoring 44.2 to 43.0 on the Noometry Index.

Last verified . 19 shared benchmarks.

Hy3 Tencent

44.2

Rank #79 Confirmed

MiMo-V2-Pro Xiaomi

43.0

Rank #103 Confirmed

Summary

  • They share 19 benchmarks with published results for both. Hy3 scores higher in 5 categories and MiMo-V2-Pro in 3 categories; 3 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Hy3 leads 26.1 to 22.1.
  • The biggest single-benchmark swing is NYT Connections (extended): 41.2% for Hy3 and 25.8% for MiMo-V2-Pro.
  • Hy3 is cheaper at $0.13 / $0.53 per million input/output tokens, against $0.43 / $0.87 for MiMo-V2-Pro.
  • MiMo-V2-Pro accepts more context: 1.05M tokens versus 262K.
  • Hy3 has downloadable open weights; the other is API-only.

Side by side

Hy3 and MiMo-V2-Pro specifications
Hy3MiMo-V2-Pro
ProviderTencentXiaomi
Noometry Index44.243.0
Released2026-07-062026-03-18
WeightsOpenProprietary
Context window262K1.05M
Max output128K131K
Input $ / M tokens$0.13$0.43
Output $ / M tokens$0.53$0.87
Results tracked1923

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hy3 leads

Hy3: 46.8 (#63), MiMo-V2-Pro: 43.8 (#83)

Coding benchmarks
BenchmarkHy3MiMo-V2-Pro
LMArena WebDev15081433
LMArena Coding14641476
ALE-Bench—785.17

Reasoning Hy3 leads

Hy3: 26.1 (#136), MiMo-V2-Pro: 22.1 (#206)

Reasoning benchmarks
BenchmarkHy3MiMo-V2-Pro
NYT Connections (extended)41.2%25.8%
LMArena Hard Prompts14471457
Thematic Generalization—45.9%

Math Too close to call

Hy3: 40.1 (#93), MiMo-V2-Pro: 39.5 (#102)

Math benchmarks
BenchmarkHy3MiMo-V2-Pro
LMArena Math14751447

Knowledge Too close to call

Hy3: 40.8 (#114), MiMo-V2-Pro: 41.4 (#111)

Knowledge benchmarks
BenchmarkHy3MiMo-V2-Pro
LMArena Expert14601478

Multilingual Too close to call

Hy3: 53.5 (#65), MiMo-V2-Pro: 52.7 (#81)

Multilingual benchmarks
BenchmarkHy3MiMo-V2-Pro
LMArena Non-English14261416
LMArena Chinese14931456
LMArena French14611469
LMArena German14391417
LMArena Japanese13921366
LMArena Korean13951400
LMArena Russian14321427
LMArena Spanish14561457

Instruction Following Too close to call

Hy3: 75.1 (#70), MiMo-V2-Pro: 76.0 (#49)

Instruction Following benchmarks
BenchmarkHy3MiMo-V2-Pro
LMArena Instruction Following14261445

Long Context Hy3 leads

Hy3: 44.1 (#75), MiMo-V2-Pro: 41.5 (#138)

Long Context benchmarks
BenchmarkHy3MiMo-V2-Pro
LMArena Longer Query14421455
CL-bench—15.7%
CL-bench Life—6.9%

Writing & Preference Too close to call

Hy3: 62.2 (#81), MiMo-V2-Pro: 62.8 (#70)

Writing & Preference benchmarks
BenchmarkHy3MiMo-V2-Pro
LMArena Text14391436
LMArena Creative Writing14021415
LMArena Multi-Turn14361456

Frequently asked questions

Is Hy3 better than MiMo-V2-Pro?

Hy3 is the stronger model overall, scoring 44.2 to 43.0 on the Noometry Index.

Which is cheaper, Hy3 or MiMo-V2-Pro?

Hy3 is cheaper. It lists at $0.13 per million input tokens and $0.53 per million output tokens; MiMo-V2-Pro lists at $0.43 and $0.87.

Is Hy3 or MiMo-V2-Pro better for coding?

Hy3 scores higher on coding benchmarks: 46.8 versus 43.8 in the Noometry coding category.

Which has the bigger context window?

MiMo-V2-Pro does, with 1.05M tokens against 262K.

How many benchmarks do Hy3 and MiMo-V2-Pro share?

19 benchmarks have published results for both models. Hy3 has 19 scored results on Noometry and MiMo-V2-Pro has 23.

Related comparisons

Go deeper