Model comparison

Gemma 3n E4b IT vs MiMo-V2.5

MiMo-V2.5 is the stronger model overall, scoring 43.4 to 37.3 on the Noometry Index.

Last verified . 17 shared benchmarks.

Gemma 3n E4b IT Google

37.3

Rank #206 Confirmed

MiMo-V2.5 Xiaomi

43.4

Rank #93 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Gemma 3n E4b IT scores higher in 0 categories and MiMo-V2.5 in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where MiMo-V2.5 leads 61.6 to 50.1.

Side by side

Gemma 3n E4b IT and MiMo-V2.5 specifications
Gemma 3n E4b ITMiMo-V2.5
ProviderGoogleXiaomi
Noometry Index37.343.4
Released—2026-04-22
WeightsOpenOpen
Context window—1.05M
Max output—131K
Input $ / M tokens—$0.14
Output $ / M tokens—$0.28
Results tracked1823

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding MiMo-V2.5 leads

Gemma 3n E4b IT: 37.0 (#198), MiMo-V2.5: 43.9 (#81)

Coding benchmarks
BenchmarkGemma 3n E4b ITMiMo-V2.5
LMArena Coding12681469
LMArena WebDev—1438
SciCode—43.1%
ALE-Bench—513.95

Reasoning MiMo-V2.5 leads

Gemma 3n E4b IT: 19.9 (#247), MiMo-V2.5: 28.6 (#101)

Reasoning benchmarks
BenchmarkGemma 3n E4b ITMiMo-V2.5
LMArena Hard Prompts12841450
Kagi LLM Benchmark31.5%—
CritPt—3.7%

Math MiMo-V2.5 leads

Gemma 3n E4b IT: 35.1 (#188), MiMo-V2.5: 36.8 (#163)

Math benchmarks
BenchmarkGemma 3n E4b ITMiMo-V2.5
LMArena Math12511436
ProofBench—16%

Knowledge MiMo-V2.5 leads

Gemma 3n E4b IT: 34.2 (#198), MiMo-V2.5: 40.8 (#115)

Knowledge benchmarks
BenchmarkGemma 3n E4b ITMiMo-V2.5
LMArena Expert12461460

Multimodal Not comparable

Gemma 3n E4b IT: —, MiMo-V2.5: 39.8 (#54)

Multimodal benchmarks
BenchmarkGemma 3n E4b ITMiMo-V2.5
LMArena Vision—1247

Multilingual MiMo-V2.5 leads

Gemma 3n E4b IT: 43.4 (#183), MiMo-V2.5: 51.9 (#99)

Multilingual benchmarks
BenchmarkGemma 3n E4b ITMiMo-V2.5
LMArena Non-English12851404
LMArena Chinese13091468
LMArena French13301447
LMArena German13111421
LMArena Japanese12721306
LMArena Korean12591363
LMArena Russian12881395
LMArena Spanish13051416

Instruction Following MiMo-V2.5 leads

Gemma 3n E4b IT: 66.1 (#210), MiMo-V2.5: 75.5 (#60)

Instruction Following benchmarks
BenchmarkGemma 3n E4b ITMiMo-V2.5
LMArena Instruction Following12551434

Long Context MiMo-V2.5 leads

Gemma 3n E4b IT: 38.7 (#191), MiMo-V2.5: 44.2 (#73)

Long Context benchmarks
BenchmarkGemma 3n E4b ITMiMo-V2.5
LMArena Longer Query12761445

Writing & Preference MiMo-V2.5 leads

Gemma 3n E4b IT: 50.1 (#186), MiMo-V2.5: 61.6 (#86)

Writing & Preference benchmarks
BenchmarkGemma 3n E4b ITMiMo-V2.5
LMArena Text13061428
LMArena Creative Writing12871393
LMArena Multi-Turn12761445

Frequently asked questions

Is Gemma 3n E4b IT better than MiMo-V2.5?

MiMo-V2.5 is the stronger model overall, scoring 43.4 to 37.3 on the Noometry Index.

Is Gemma 3n E4b IT or MiMo-V2.5 better for coding?

MiMo-V2.5 scores higher on coding benchmarks: 43.9 versus 37.0 in the Noometry coding category.

How many benchmarks do Gemma 3n E4b IT and MiMo-V2.5 share?

17 benchmarks have published results for both models. Gemma 3n E4b IT has 18 scored results on Noometry and MiMo-V2.5 has 23.

Related comparisons

Go deeper