Model comparison

Chatgpt 4o Latest 20250326 vs MiMo-V2.5

Chatgpt 4o Latest 20250326 and MiMo-V2.5 score almost the same on the Noometry Index (43.8 vs 43.4), so choose on price, context window or the category you care about most.

Last verified . 18 shared benchmarks.

Chatgpt 4o Latest 20250326 OpenAI

43.8

Rank #82 Confirmed

MiMo-V2.5 Xiaomi

43.4

Rank #93 Confirmed

Summary

  • They share 18 benchmarks with published results for both. Chatgpt 4o Latest 20250326 scores higher in 4 categories and MiMo-V2.5 in 5 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Chatgpt 4o Latest 20250326 leads 33.8 to 28.6.
  • MiMo-V2.5 has downloadable open weights; the other is API-only.

Side by side

Chatgpt 4o Latest 20250326 and MiMo-V2.5 specifications
Chatgpt 4o Latest 20250326MiMo-V2.5
ProviderOpenAIXiaomi
Noometry Index43.843.4
Released—2026-04-22
WeightsProprietaryOpen
Context window—1.05M
Max output—131K
Input $ / M tokens—$0.14
Output $ / M tokens—$0.28
Results tracked2123

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding MiMo-V2.5 leads

Chatgpt 4o Latest 20250326: 41.6 (#122), MiMo-V2.5: 43.9 (#81)

Coding benchmarks
BenchmarkChatgpt 4o Latest 20250326MiMo-V2.5
LMArena Coding14131469
LMArena WebDev—1438
SciCode—43.1%
ALE-Bench—513.95

Reasoning Chatgpt 4o Latest 20250326 leads

Chatgpt 4o Latest 20250326: 33.8 (#71), MiMo-V2.5: 28.6 (#101)

Reasoning benchmarks
BenchmarkChatgpt 4o Latest 20250326MiMo-V2.5
LMArena Hard Prompts14241450
Kagi LLM Benchmark75%—
CritPt—3.7%

Math Chatgpt 4o Latest 20250326 leads

Chatgpt 4o Latest 20250326: 38.6 (#134), MiMo-V2.5: 36.8 (#163)

Math benchmarks
BenchmarkChatgpt 4o Latest 20250326MiMo-V2.5
LMArena Math14071436
ProofBench—16%

Knowledge MiMo-V2.5 leads

Chatgpt 4o Latest 20250326: 39.2 (#137), MiMo-V2.5: 40.8 (#115)

Knowledge benchmarks
BenchmarkChatgpt 4o Latest 20250326MiMo-V2.5
LMArena Expert14011460
Confabulations16.6%—

Multimodal Too close to call

Chatgpt 4o Latest 20250326: 39.6 (#58), MiMo-V2.5: 39.8 (#54)

Multimodal benchmarks
BenchmarkChatgpt 4o Latest 20250326MiMo-V2.5
LMArena Vision12431247

Multilingual Chatgpt 4o Latest 20250326 leads

Chatgpt 4o Latest 20250326: 52.9 (#76), MiMo-V2.5: 51.9 (#99)

Multilingual benchmarks
BenchmarkChatgpt 4o Latest 20250326MiMo-V2.5
LMArena Non-English14191404
LMArena Chinese14571468
LMArena French14461447
LMArena German14231421
LMArena Japanese14051306
LMArena Korean13961363
LMArena Russian14291395
LMArena Spanish14341416

Instruction Following MiMo-V2.5 leads

Chatgpt 4o Latest 20250326: 74.0 (#107), MiMo-V2.5: 75.5 (#60)

Instruction Following benchmarks
BenchmarkChatgpt 4o Latest 20250326MiMo-V2.5
LMArena Instruction Following14031434

Long Context MiMo-V2.5 leads

Chatgpt 4o Latest 20250326: 43.1 (#107), MiMo-V2.5: 44.2 (#73)

Long Context benchmarks
BenchmarkChatgpt 4o Latest 20250326MiMo-V2.5
LMArena Longer Query14131445

Writing & Preference Chatgpt 4o Latest 20250326 leads

Chatgpt 4o Latest 20250326: 62.6 (#72), MiMo-V2.5: 61.6 (#86)

Writing & Preference benchmarks
BenchmarkChatgpt 4o Latest 20250326MiMo-V2.5
LMArena Text14291428
LMArena Creative Writing14051393
LMArena Multi-Turn14541445
EQ-Bench Creative Writing1501—

Frequently asked questions

Is Chatgpt 4o Latest 20250326 better than MiMo-V2.5?

Chatgpt 4o Latest 20250326 and MiMo-V2.5 score almost the same on the Noometry Index (43.8 vs 43.4), so choose on price, context window or the category you care about most.

Is Chatgpt 4o Latest 20250326 or MiMo-V2.5 better for coding?

MiMo-V2.5 scores higher on coding benchmarks: 43.9 versus 41.6 in the Noometry coding category.

How many benchmarks do Chatgpt 4o Latest 20250326 and MiMo-V2.5 share?

18 benchmarks have published results for both models. Chatgpt 4o Latest 20250326 has 21 scored results on Noometry and MiMo-V2.5 has 23.

Related comparisons

Go deeper