Model comparison

Gemini 3.1 Flash Lite vs MiMo-V2.5

MiMo-V2.5 is the stronger model overall, scoring 43.4 to 40.8 on the Noometry Index.

Last verified . 22 shared benchmarks.

Gemini 3.1 Flash Lite Google

40.8

Rank #144 Confirmed

MiMo-V2.5 Xiaomi

43.4

Rank #93 Confirmed

Summary

  • They share 22 benchmarks with published results for both. Gemini 3.1 Flash Lite scores higher in 3 categories and MiMo-V2.5 in 6 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in coding, where MiMo-V2.5 leads 43.9 to 37.8.
  • MiMo-V2.5 is cheaper at $0.14 / $0.28 per million input/output tokens, against $0.25 / $1.50 for Gemini 3.1 Flash Lite.
  • MiMo-V2.5 has downloadable open weights; the other is API-only.

Side by side

Gemini 3.1 Flash Lite and MiMo-V2.5 specifications
Gemini 3.1 Flash LiteMiMo-V2.5
ProviderGoogleXiaomi
Noometry Index40.843.4
Released2026-03-032026-04-22
WeightsProprietaryOpen
Context window1.05M1.05M
Max output66K131K
Input $ / M tokens$0.25$0.14
Output $ / M tokens$1.50$0.28
Results tracked3823

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding MiMo-V2.5 leads

Gemini 3.1 Flash Lite: 37.8 (#188), MiMo-V2.5: 43.9 (#81)

Coding benchmarks
BenchmarkGemini 3.1 Flash LiteMiMo-V2.5
LMArena WebDev12561438
SciCode41.9%43.1%
LMArena Coding14001469
ALE-Bench797.73513.95
WeirdML52.2%—

Agentic & Tool Use Not comparable

Gemini 3.1 Flash Lite: 30.2 (#79), MiMo-V2.5: —

Agentic & Tool Use benchmarks
BenchmarkGemini 3.1 Flash LiteMiMo-V2.5
DeepResearch Bench37.3%—

Reasoning MiMo-V2.5 leads

Gemini 3.1 Flash Lite: 22.9 (#186), MiMo-V2.5: 28.6 (#101)

Reasoning benchmarks
BenchmarkGemini 3.1 Flash LiteMiMo-V2.5
CritPt1.1%3.7%
LMArena Hard Prompts14071450
Kagi LLM Benchmark67.2%—
NYT Connections (extended)8.2%—
Chess Puzzles25%—
EnigmaEval3%—
Thematic Generalization63.3%—
DTBench76.8%—
LMCA35%—
Epoch Capabilities Index144.47—
ForecastBench54.4—

Math Gemini 3.1 Flash Lite leads

Gemini 3.1 Flash Lite: 40.7 (#90), MiMo-V2.5: 36.8 (#163)

Math benchmarks
BenchmarkGemini 3.1 Flash LiteMiMo-V2.5
LMArena Math14281436
FrontierMath (Tiers 1-3)27.7%—
OTIS Mock AIME 2024-202580%—
ProofBench—16%

Knowledge Gemini 3.1 Flash Lite leads

Gemini 3.1 Flash Lite: 41.9 (#104), MiMo-V2.5: 40.8 (#115)

Knowledge benchmarks
BenchmarkGemini 3.1 Flash LiteMiMo-V2.5
LMArena Expert13981460
GPQA Diamond81.8%—
Humanity's Last Exam8.6%—
Vectara Hallucination Rate8.2%—

Multimodal Too close to call

Gemini 3.1 Flash Lite: 39.4 (#60), MiMo-V2.5: 39.8 (#54)

Multimodal benchmarks
BenchmarkGemini 3.1 Flash LiteMiMo-V2.5
LMArena Vision12401247

Multilingual Too close to call

Gemini 3.1 Flash Lite: 52.3 (#86), MiMo-V2.5: 51.9 (#99)

Multilingual benchmarks
BenchmarkGemini 3.1 Flash LiteMiMo-V2.5
LMArena Non-English14111404
LMArena Chinese14611468
LMArena French14241447
LMArena German14291421
LMArena Japanese14131306
LMArena Korean13921363
LMArena Russian14201395
LMArena Spanish14211416

Instruction Following MiMo-V2.5 leads

Gemini 3.1 Flash Lite: 72.7 (#131), MiMo-V2.5: 75.5 (#60)

Instruction Following benchmarks
BenchmarkGemini 3.1 Flash LiteMiMo-V2.5
LMArena Instruction Following13771434

Long Context MiMo-V2.5 leads

Gemini 3.1 Flash Lite: 42.5 (#122), MiMo-V2.5: 44.2 (#73)

Long Context benchmarks
BenchmarkGemini 3.1 Flash LiteMiMo-V2.5
LMArena Longer Query13941445

Writing & Preference Too close to call

Gemini 3.1 Flash Lite: 60.9 (#94), MiMo-V2.5: 61.6 (#86)

Writing & Preference benchmarks
BenchmarkGemini 3.1 Flash LiteMiMo-V2.5
LMArena Text14161428
LMArena Creative Writing14011393
LMArena Multi-Turn14171445

Frequently asked questions

Is Gemini 3.1 Flash Lite better than MiMo-V2.5?

MiMo-V2.5 is the stronger model overall, scoring 43.4 to 40.8 on the Noometry Index.

Which is cheaper, Gemini 3.1 Flash Lite or MiMo-V2.5?

MiMo-V2.5 is cheaper. It lists at $0.14 per million input tokens and $0.28 per million output tokens; Gemini 3.1 Flash Lite lists at $0.25 and $1.50.

Is Gemini 3.1 Flash Lite or MiMo-V2.5 better for coding?

MiMo-V2.5 scores higher on coding benchmarks: 43.9 versus 37.8 in the Noometry coding category.

Which has the bigger context window?

Both accept 1.05M tokens.

How many benchmarks do Gemini 3.1 Flash Lite and MiMo-V2.5 share?

22 benchmarks have published results for both models. Gemini 3.1 Flash Lite has 38 scored results on Noometry and MiMo-V2.5 has 23.

Related comparisons

Go deeper