Model comparison

Gemma 3 27B vs MiMo-V2.5-Pro

MiMo-V2.5-Pro is the stronger model overall, scoring 45.2 to 30.8 on the Noometry Index. Gemma 3 27B costs 5.4× less per token, which makes it the better buy when MiMo-V2.5-Pro's lead doesn't matter for your workload.

Last verified . 22 shared benchmarks.

Gemma 3 27B Google

30.8

Rank #284 Confirmed

MiMo-V2.5-Pro Xiaomi

45.2

Rank #74 Confirmed

Summary

  • They share 22 benchmarks with published results for both. Gemma 3 27B scores higher in 0 categories and MiMo-V2.5-Pro in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in coding, where MiMo-V2.5-Pro leads 47.4 to 22.5.
  • The biggest single-benchmark swing is DTBench: 52.5% for Gemma 3 27B and 84.5% for MiMo-V2.5-Pro.
  • Gemma 3 27B is cheaper at $0.08 / $0.16 per million input/output tokens, against $0.43 / $0.87 for MiMo-V2.5-Pro.
  • MiMo-V2.5-Pro accepts more context: 1.05M tokens versus 131K.

Side by side

Gemma 3 27B and MiMo-V2.5-Pro specifications
Gemma 3 27BMiMo-V2.5-Pro
ProviderGoogleXiaomi
Noometry Index30.845.2
Released2025-03-112026-04-22
WeightsOpenOpen
Context window131K1.05M
Max output8K131K
Input $ / M tokens$0.08$0.43
Output $ / M tokens$0.16$0.87
Results tracked4327

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding MiMo-V2.5-Pro leads

Gemma 3 27B: 22.5 (#334), MiMo-V2.5-Pro: 47.4 (#60)

Coding benchmarks
BenchmarkGemma 3 27BMiMo-V2.5-Pro
SciCode21.2%50.2%
LMArena Coding13221503
Aider Polyglot4.9%—
LMArena WebDev—1479
LiveBench Coding39.9%—
ALE-Bench—899.8

Agentic & Tool Use Not comparable

Gemma 3 27B: 25.1 (#110), MiMo-V2.5-Pro: —

Agentic & Tool Use benchmarks
BenchmarkGemma 3 27BMiMo-V2.5-Pro
Berkeley Function Calling Leaderboard29.5%—

Reasoning MiMo-V2.5-Pro leads

Gemma 3 27B: 16.7 (#301), MiMo-V2.5-Pro: 26.8 (#130)

Reasoning benchmarks
BenchmarkGemma 3 27BMiMo-V2.5-Pro
CritPt0%4%
LMArena Hard Prompts13401488
DTBench52.5%84.5%
LMCA12.3%29.5%
Kagi LLM Benchmark40.4%—
NYT Connections (extended)—34.4%
Chess Puzzles0%—
LiveBench Reasoning43.8%—
LiveBench Data Analysis51.5%—
Epoch Capabilities Index130.04—
LiveBench50%—

Math MiMo-V2.5-Pro leads

Gemma 3 27B: 25.9 (#265), MiMo-V2.5-Pro: 40.0 (#96)

Math benchmarks
BenchmarkGemma 3 27BMiMo-V2.5-Pro
LMArena Math13121481
OTIS Mock AIME 2024-202522.5%—
ProofBench—22%
LiveBench Math55.4%—
MATH Level 574%—

Knowledge MiMo-V2.5-Pro leads

Gemma 3 27B: 25.5 (#261), MiMo-V2.5-Pro: 42.2 (#98)

Knowledge benchmarks
BenchmarkGemma 3 27BMiMo-V2.5-Pro
LMArena Expert13041503
GPQA Diamond47.7%—
Confabulations40.3%—
Vectara Hallucination Rate7.4%—

Multimodal Not comparable

Gemma 3 27B: 32.6 (#100), MiMo-V2.5-Pro: —

Multimodal benchmarks
BenchmarkGemma 3 27BMiMo-V2.5-Pro
LMArena Vision1164—
GeoBench52%—

Multilingual MiMo-V2.5-Pro leads

Gemma 3 27B: 46.9 (#155), MiMo-V2.5-Pro: 55.1 (#34)

Multilingual benchmarks
BenchmarkGemma 3 27BMiMo-V2.5-Pro
LMArena Non-English13341449
LMArena Chinese13461507
LMArena French13681488
LMArena German13621458
LMArena Japanese12871412
LMArena Korean13081437
LMArena Russian13491450
LMArena Spanish13491471

Instruction Following MiMo-V2.5-Pro leads

Gemma 3 27B: 70.6 (#160), MiMo-V2.5-Pro: 77.5 (#21)

Instruction Following benchmarks
BenchmarkGemma 3 27BMiMo-V2.5-Pro
LMArena Instruction Following13211477
LiveBench Instruction Following74.9%—

Long Context MiMo-V2.5-Pro leads

Gemma 3 27B: 27.6 (#293), MiMo-V2.5-Pro: 45.4 (#37)

Long Context benchmarks
BenchmarkGemma 3 27BMiMo-V2.5-Pro
LMArena Longer Query13331483
Fiction.LiveBench33.3%—

Writing & Preference MiMo-V2.5-Pro leads

Gemma 3 27B: 52.5 (#168), MiMo-V2.5-Pro: 65.3 (#49)

Writing & Preference benchmarks
BenchmarkGemma 3 27BMiMo-V2.5-Pro
LMArena Text13581465
LMArena Creative Writing13461440
EQ-Bench Creative Writing12661493
LMArena Multi-Turn13451477
Short-Story Creative Writing79.9%—
EQ-Bench 4—1208
LiveBench Language34.6%—

Frequently asked questions

Is Gemma 3 27B better than MiMo-V2.5-Pro?

MiMo-V2.5-Pro is the stronger model overall, scoring 45.2 to 30.8 on the Noometry Index. Gemma 3 27B costs 5.4× less per token, which makes it the better buy when MiMo-V2.5-Pro's lead doesn't matter for your workload.

Which is cheaper, Gemma 3 27B or MiMo-V2.5-Pro?

Gemma 3 27B is cheaper. It lists at $0.08 per million input tokens and $0.16 per million output tokens; MiMo-V2.5-Pro lists at $0.43 and $0.87.

Is Gemma 3 27B or MiMo-V2.5-Pro better for coding?

MiMo-V2.5-Pro scores higher on coding benchmarks: 47.4 versus 22.5 in the Noometry coding category.

Which has the bigger context window?

MiMo-V2.5-Pro does, with 1.05M tokens against 131K.

How many benchmarks do Gemma 3 27B and MiMo-V2.5-Pro share?

22 benchmarks have published results for both models. Gemma 3 27B has 43 scored results on Noometry and MiMo-V2.5-Pro has 27.

Related comparisons

Go deeper