Model comparison

Granite 4.0 H Small vs MiniMax-M2.7

MiniMax-M2.7 is the stronger model overall, scoring 37.7 to 36.5 on the Noometry Index.

Last verified . 14 shared benchmarks.

Granite 4.0 H Small IBM

36.5

Rank #214 Confirmed

MiniMax-M2.7 MiniMax

37.7

Rank #196 Confirmed

Summary

  • They share 14 benchmarks with published results for both. Granite 4.0 H Small scores higher in 2 categories and MiniMax-M2.7 in 6 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where MiniMax-M2.7 leads 58.9 to 43.8.
  • The biggest single-benchmark swing is Vectara Hallucination Rate: 5.2% for Granite 4.0 H Small and 12.9% for MiniMax-M2.7.

Side by side

Granite 4.0 H Small and MiniMax-M2.7 specifications
Granite 4.0 H SmallMiniMax-M2.7
ProviderIBMMiniMax
Noometry Index36.537.7
Released—2026-03-18
WeightsOpenOpen
Context window—205K
Max output—131K
Input $ / M tokens—$0.30
Output $ / M tokens—$1.20
Results tracked1930

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding MiniMax-M2.7 leads

Granite 4.0 H Small: 36.4 (#209), MiniMax-M2.7: 41.8 (#120)

Coding benchmarks
BenchmarkGranite 4.0 H SmallMiniMax-M2.7
LMArena Coding12491454
LMArena WebDev—1398
SciCode—47%
WeirdML—37%
ALE-Bench—599.25

Agentic & Tool Use Not comparable

Granite 4.0 H Small: —, MiniMax-M2.7: 25.1 (#111)

Agentic & Tool Use benchmarks
BenchmarkGranite 4.0 H SmallMiniMax-M2.7
Terminal-Bench—45.1%
ExploitBench—13.3%
GBAEval—0%

Reasoning Granite 4.0 H Small leads

Granite 4.0 H Small: 24.4 (#163), MiniMax-M2.7: 19.7 (#253)

Reasoning benchmarks
BenchmarkGranite 4.0 H SmallMiniMax-M2.7
LMArena Hard Prompts12401422
NYT Connections (extended)—24.7%
CritPt—0.6%
Thematic Generalization—39.3%
Epoch Capabilities Index—145.85

Math Granite 4.0 H Small leads

Granite 4.0 H Small: 31.4 (#223), MiniMax-M2.7: 25.9 (#263)

Math benchmarks
BenchmarkGranite 4.0 H SmallMiniMax-M2.7
LMArena Math12471420
ProofBench—3%
Omni-MATH29.6%—

Knowledge MiniMax-M2.7 leads

Granite 4.0 H Small: 32.0 (#215), MiniMax-M2.7: 37.7 (#152)

Knowledge benchmarks
BenchmarkGranite 4.0 H SmallMiniMax-M2.7
Vectara Hallucination Rate5.2%12.9%
LMArena Expert12511444
MMLU-Pro56.9%—
GPQA (HELM)38.3%—

Multilingual MiniMax-M2.7 leads

Granite 4.0 H Small: 38.6 (#228), MiniMax-M2.7: 50.3 (#123)

Multilingual benchmarks
BenchmarkGranite 4.0 H SmallMiniMax-M2.7
LMArena Non-English12161382
LMArena Chinese12491441
LMArena Russian12011383
LMArena Spanish12581403
LMArena French—1421
LMArena German—1398
LMArena Japanese—1262
LMArena Korean—1313

Instruction Following MiniMax-M2.7 leads

Granite 4.0 H Small: 68.7 (#183), MiniMax-M2.7: 74.1 (#103)

Instruction Following benchmarks
BenchmarkGranite 4.0 H SmallMiniMax-M2.7
LMArena Instruction Following12221405
IFEval89%—

Long Context MiniMax-M2.7 leads

Granite 4.0 H Small: 37.7 (#213), MiniMax-M2.7: 43.3 (#99)

Long Context benchmarks
BenchmarkGranite 4.0 H SmallMiniMax-M2.7
LMArena Longer Query12421419

Writing & Preference MiniMax-M2.7 leads

Granite 4.0 H Small: 43.8 (#227), MiniMax-M2.7: 58.9 (#112)

Writing & Preference benchmarks
BenchmarkGranite 4.0 H SmallMiniMax-M2.7
LMArena Text12411405
LMArena Creative Writing12111354
LMArena Multi-Turn12421412
WildBench73.9%—

Frequently asked questions

Is Granite 4.0 H Small better than MiniMax-M2.7?

MiniMax-M2.7 is the stronger model overall, scoring 37.7 to 36.5 on the Noometry Index.

Is Granite 4.0 H Small or MiniMax-M2.7 better for coding?

MiniMax-M2.7 scores higher on coding benchmarks: 41.8 versus 36.4 in the Noometry coding category.

How many benchmarks do Granite 4.0 H Small and MiniMax-M2.7 share?

14 benchmarks have published results for both models. Granite 4.0 H Small has 19 scored results on Noometry and MiniMax-M2.7 has 30.

Related comparisons

Go deeper