Model comparison

MiniMax-M2.7 vs Qwen3.5 Max Preview

Qwen3.5 Max Preview is the stronger model overall, scoring 45.3 to 37.7 on the Noometry Index.

Last verified . 17 shared benchmarks.

MiniMax-M2.7 MiniMax

37.7

Rank #196 Confirmed

Qwen3.5 Max Preview Alibaba (Qwen)

45.3

Rank #71 Confirmed

Summary

  • They share 17 benchmarks with published results for both. MiniMax-M2.7 scores higher in 0 categories and Qwen3.5 Max Preview in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in math, where Qwen3.5 Max Preview leads 40.1 to 25.9.
  • MiniMax-M2.7 has downloadable open weights; the other is API-only.

Side by side

MiniMax-M2.7 and Qwen3.5 Max Preview specifications
MiniMax-M2.7Qwen3.5 Max Preview
ProviderMiniMaxAlibaba (Qwen)
Noometry Index37.745.3
Released2026-03-18—
WeightsOpenProprietary
Context window205K—
Max output131K—
Input $ / M tokens$0.30—
Output $ / M tokens$1.20—
Results tracked3017

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen3.5 Max Preview leads

MiniMax-M2.7: 41.8 (#120), Qwen3.5 Max Preview: 44.0 (#77)

Coding benchmarks
BenchmarkMiniMax-M2.7Qwen3.5 Max Preview
LMArena Coding14541487
LMArena WebDev1398—
SciCode47%—
WeirdML37%—
ALE-Bench599.25—

Agentic & Tool Use Not comparable

MiniMax-M2.7: 25.1 (#111), Qwen3.5 Max Preview: —

Agentic & Tool Use benchmarks
BenchmarkMiniMax-M2.7Qwen3.5 Max Preview
Terminal-Bench45.1%—
ExploitBench13.3%—
GBAEval0%—

Reasoning Qwen3.5 Max Preview leads

MiniMax-M2.7: 19.7 (#253), Qwen3.5 Max Preview: 30.8 (#84)

Reasoning benchmarks
BenchmarkMiniMax-M2.7Qwen3.5 Max Preview
LMArena Hard Prompts14221483
NYT Connections (extended)24.7%—
CritPt0.6%—
Thematic Generalization39.3%—
Epoch Capabilities Index145.85—

Math Qwen3.5 Max Preview leads

MiniMax-M2.7: 25.9 (#263), Qwen3.5 Max Preview: 40.1 (#94)

Math benchmarks
BenchmarkMiniMax-M2.7Qwen3.5 Max Preview
LMArena Math14201474
ProofBench3%—

Knowledge Qwen3.5 Max Preview leads

MiniMax-M2.7: 37.7 (#152), Qwen3.5 Max Preview: 41.8 (#107)

Knowledge benchmarks
BenchmarkMiniMax-M2.7Qwen3.5 Max Preview
LMArena Expert14441489
Vectara Hallucination Rate12.9%—

Multilingual Qwen3.5 Max Preview leads

MiniMax-M2.7: 50.3 (#123), Qwen3.5 Max Preview: 56.2 (#22)

Multilingual benchmarks
BenchmarkMiniMax-M2.7Qwen3.5 Max Preview
LMArena Non-English13821465
LMArena Chinese14411534
LMArena French14211484
LMArena German13981487
LMArena Japanese12621495
LMArena Korean13131438
LMArena Russian13831471
LMArena Spanish14031470

Instruction Following Qwen3.5 Max Preview leads

MiniMax-M2.7: 74.1 (#103), Qwen3.5 Max Preview: 77.0 (#31)

Instruction Following benchmarks
BenchmarkMiniMax-M2.7Qwen3.5 Max Preview
LMArena Instruction Following14051467

Long Context Qwen3.5 Max Preview leads

MiniMax-M2.7: 43.3 (#99), Qwen3.5 Max Preview: 45.2 (#45)

Long Context benchmarks
BenchmarkMiniMax-M2.7Qwen3.5 Max Preview
LMArena Longer Query14191476

Writing & Preference Qwen3.5 Max Preview leads

MiniMax-M2.7: 58.9 (#112), Qwen3.5 Max Preview: 66.0 (#41)

Writing & Preference benchmarks
BenchmarkMiniMax-M2.7Qwen3.5 Max Preview
LMArena Text14051470
LMArena Creative Writing13541464
LMArena Multi-Turn14121478

Frequently asked questions

Is MiniMax-M2.7 better than Qwen3.5 Max Preview?

Qwen3.5 Max Preview is the stronger model overall, scoring 45.3 to 37.7 on the Noometry Index.

Is MiniMax-M2.7 or Qwen3.5 Max Preview better for coding?

Qwen3.5 Max Preview scores higher on coding benchmarks: 44.0 versus 41.8 in the Noometry coding category.

How many benchmarks do MiniMax-M2.7 and Qwen3.5 Max Preview share?

17 benchmarks have published results for both models. MiniMax-M2.7 has 30 scored results on Noometry and Qwen3.5 Max Preview has 17.

Related comparisons

Go deeper