Model comparison

MiMo-V2.5 vs Qwen3.5 Plus

MiMo-V2.5 and Qwen3.5 Plus score almost the same on the Noometry Index (43.4 vs 42.9), so choose on price, context window or the category you care about most.

Last verified . 1 shared benchmarks.

MiMo-V2.5 Xiaomi

43.4

Rank #93 Confirmed

Qwen3.5 Plus Alibaba (Qwen)

42.9

Rank #106 Confirmed

Summary

  • They share 1 benchmark with published results for both. MiMo-V2.5 scores higher in 1 category and Qwen3.5 Plus in 3 categories; 4 gaps are clear of the uncertainty.
  • The widest gap is in math, where Qwen3.5 Plus leads 49.6 to 36.8.
  • MiMo-V2.5 is cheaper at $0.14 / $0.28 per million input/output tokens, against $0.40 / $2.40 for Qwen3.5 Plus.
  • MiMo-V2.5 accepts more context: 1.05M tokens versus 1M.
  • MiMo-V2.5 has downloadable open weights; the other is API-only.

Side by side

MiMo-V2.5 and Qwen3.5 Plus specifications
MiMo-V2.5Qwen3.5 Plus
ProviderXiaomiAlibaba (Qwen)
Noometry Index43.442.9
Released2026-04-222026-02-16
WeightsOpenProprietary
Context window1.05M1M
Max output131K66K
Input $ / M tokens$0.14$0.40
Output $ / M tokens$0.28$2.40
Results tracked2315

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

MiMo-V2.5: 43.9 (#81), Qwen3.5 Plus: —

Coding benchmarks
BenchmarkMiMo-V2.5Qwen3.5 Plus
ALE-Bench513.95621.92
LMArena WebDev1438—
SciCode43.1%—
LMArena Coding1469—

Agentic & Tool Use Not comparable

MiMo-V2.5: —, Qwen3.5 Plus: —

Agentic & Tool Use benchmarks
BenchmarkMiMo-V2.5Qwen3.5 Plus
Vending-Bench 2—0.54

Reasoning Qwen3.5 Plus leads

MiMo-V2.5: 28.6 (#101), Qwen3.5 Plus: 32.8 (#74)

Reasoning benchmarks
BenchmarkMiMo-V2.5Qwen3.5 Plus
CritPt3.7%—
Chess Puzzles—22%
LMArena Hard Prompts1450—
Mystery Game Puzzles—17%
DTBench—80.5%
LMCA—36.4%
Epoch Capabilities Index—146.78

Math Qwen3.5 Plus leads

MiMo-V2.5: 36.8 (#163), Qwen3.5 Plus: 49.6 (#61)

Math benchmarks
BenchmarkMiMo-V2.5Qwen3.5 Plus
OTIS Mock AIME 2024-2025—86.7%
ProofBench16%—
LMArena Math1436—
FrontierMath (Feb 2025 set)—21%
FrontierMath Tier 4 (v1)—2.1%

Knowledge Qwen3.5 Plus leads

MiMo-V2.5: 40.8 (#115), Qwen3.5 Plus: 46.0 (#83)

Knowledge benchmarks
BenchmarkMiMo-V2.5Qwen3.5 Plus
GPQA Diamond—84.8%
SimpleQA Verified—25.4%
Vectara Hallucination Rate—10.7%
LMArena Expert1460—

Multimodal Not comparable

MiMo-V2.5: 39.8 (#54), Qwen3.5 Plus: —

Multimodal benchmarks
BenchmarkMiMo-V2.5Qwen3.5 Plus
LMArena Vision1247—

Multilingual Not comparable

MiMo-V2.5: 51.9 (#99), Qwen3.5 Plus: —

Multilingual benchmarks
BenchmarkMiMo-V2.5Qwen3.5 Plus
LMArena Non-English1404—
LMArena Chinese1468—
LMArena French1447—
LMArena German1421—
LMArena Japanese1306—
LMArena Korean1363—
LMArena Russian1395—
LMArena Spanish1416—

Instruction Following Not comparable

MiMo-V2.5: 75.5 (#60), Qwen3.5 Plus: —

Instruction Following benchmarks
BenchmarkMiMo-V2.5Qwen3.5 Plus
LMArena Instruction Following1434—

Long Context MiMo-V2.5 leads

MiMo-V2.5: 44.2 (#73), Qwen3.5 Plus: 43.0 (#113)

Long Context benchmarks
BenchmarkMiMo-V2.5Qwen3.5 Plus
CL-bench—19.8%
CL-bench Life—12.4%
LMArena Longer Query1445—

Writing & Preference Not comparable

MiMo-V2.5: 61.6 (#86), Qwen3.5 Plus: —

Writing & Preference benchmarks
BenchmarkMiMo-V2.5Qwen3.5 Plus
LMArena Text1428—
LMArena Creative Writing1393—
LMArena Multi-Turn1445—

Frequently asked questions

Is MiMo-V2.5 better than Qwen3.5 Plus?

MiMo-V2.5 and Qwen3.5 Plus score almost the same on the Noometry Index (43.4 vs 42.9), so choose on price, context window or the category you care about most.

Which is cheaper, MiMo-V2.5 or Qwen3.5 Plus?

MiMo-V2.5 is cheaper. It lists at $0.14 per million input tokens and $0.28 per million output tokens; Qwen3.5 Plus lists at $0.40 and $2.40.

Which has the bigger context window?

MiMo-V2.5 does, with 1.05M tokens against 1M.

How many benchmarks do MiMo-V2.5 and Qwen3.5 Plus share?

1 benchmark has published results for both models. MiMo-V2.5 has 23 scored results on Noometry and Qwen3.5 Plus has 15.

Related comparisons

Go deeper