Model comparison

Granite 4.0 Micro vs Granite 4.0 H Small

Granite 4.0 H Small is the stronger model overall, scoring 36.5 to 29.0 on the Noometry Index.

Last verified . 5 shared benchmarks.

Granite 4.0 Micro IBM

29.0

Rank #318 Confirmed

Granite 4.0 H Small IBM

36.5

Rank #214 Confirmed

Summary

  • They share 5 benchmarks with published results for both. Granite 4.0 Micro scores higher in 2 categories and Granite 4.0 H Small in 3 categories; 5 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Granite 4.0 H Small leads 32.0 to 9.9.
  • The biggest single-benchmark swing is MMLU-Pro: 39.5% for Granite 4.0 Micro and 56.9% for Granite 4.0 H Small.

Side by side

Granite 4.0 Micro and Granite 4.0 H Small specifications
Granite 4.0 MicroGranite 4.0 H Small
ProviderIBMIBM
Noometry Index29.036.5
Released2025-10-02—
WeightsOpenOpen
Context window131K—
Max output118K—
Input $ / M tokens$0.017—
Output $ / M tokens$0.11—
Results tracked819

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Granite 4.0 Micro: —, Granite 4.0 H Small: 36.4 (#209)

Coding benchmarks
BenchmarkGranite 4.0 MicroGranite 4.0 H Small
LMArena Coding—1249

Reasoning Granite 4.0 H Small leads

Granite 4.0 Micro: 19.2 (#265), Granite 4.0 H Small: 24.4 (#163)

Reasoning benchmarks
BenchmarkGranite 4.0 MicroGranite 4.0 H Small
Chess Puzzles0%—
LMArena Hard Prompts—1240

Math Granite 4.0 H Small leads

Granite 4.0 Micro: 12.0 (#307), Granite 4.0 H Small: 31.4 (#223)

Math benchmarks
BenchmarkGranite 4.0 MicroGranite 4.0 H Small
Omni-MATH20.9%29.6%
OTIS Mock AIME 2024-20252.8%—
LMArena Math—1247

Knowledge Granite 4.0 H Small leads

Granite 4.0 Micro: 9.9 (#304), Granite 4.0 H Small: 32.0 (#215)

Knowledge benchmarks
BenchmarkGranite 4.0 MicroGranite 4.0 H Small
MMLU-Pro39.5%56.9%
GPQA (HELM)30.7%38.3%
GPQA Diamond28.3%—
Vectara Hallucination Rate—5.2%
LMArena Expert—1251

Multilingual Not comparable

Granite 4.0 Micro: —, Granite 4.0 H Small: 38.6 (#228)

Multilingual benchmarks
BenchmarkGranite 4.0 MicroGranite 4.0 H Small
LMArena Non-English—1216
LMArena Chinese—1249
LMArena Russian—1201
LMArena Spanish—1258

Instruction Following Granite 4.0 Micro leads

Granite 4.0 Micro: 69.9 (#169), Granite 4.0 H Small: 68.7 (#183)

Instruction Following benchmarks
BenchmarkGranite 4.0 MicroGranite 4.0 H Small
IFEval84.9%89%
LMArena Instruction Following—1222

Long Context Not comparable

Granite 4.0 Micro: —, Granite 4.0 H Small: 37.7 (#213)

Long Context benchmarks
BenchmarkGranite 4.0 MicroGranite 4.0 H Small
LMArena Longer Query—1242

Writing & Preference Granite 4.0 Micro leads

Granite 4.0 Micro: 46.7 (#216), Granite 4.0 H Small: 43.8 (#227)

Writing & Preference benchmarks
BenchmarkGranite 4.0 MicroGranite 4.0 H Small
WildBench67%73.9%
LMArena Text—1241
LMArena Creative Writing—1211
LMArena Multi-Turn—1242

Frequently asked questions

Is Granite 4.0 Micro better than Granite 4.0 H Small?

Granite 4.0 H Small is the stronger model overall, scoring 36.5 to 29.0 on the Noometry Index.

How many benchmarks do Granite 4.0 Micro and Granite 4.0 H Small share?

5 benchmarks have published results for both models. Granite 4.0 Micro has 8 scored results on Noometry and Granite 4.0 H Small has 19.

Related comparisons

Go deeper