IBM, open weights

Granite 4.0 H Small

Granite 4.0 H Small by IBM ranks 214th of 354 ranked models on the Noometry Index as of October 2026, with a score of 36.5. Its strongest category is reasoning, where it ranks 163rd.

Last verified

Specifications

Noometry rank
#214 of 354
Index score
36.5
Evidence
Confirmed 19 results
Provider
IBM
Released
Unknown
Weights
Open weights
Reasoning
Unknown
Context window
—
Max output
—
Input price
Not listed
Output price
Not listed
Blended price
Not listed
Output speed
Not measured
Value
Not ranked
Knowledge cutoff
Unknown

Category scores

Each category score combines every public result we have in that category.

Granite 4.0 H Small category scores
  1. Coding 36.4
  2. Reasoning 24.4
  3. Math 31.4
  4. Knowledge 32.0
  5. Multilingual 38.6
  6. Instruction Following 68.7
  7. Long Context 37.7
  8. Writing & Preference 43.8
Granite 4.0 H Small category ranks
CategoryScoreRankResults
Coding36.4#2091
Reasoning24.4#1631
Math31.4#2232
Knowledge32.0#2154
Multilingual38.6#2281
Instruction Following68.7#1832
Long Context37.7#2131
Writing & Preference43.8#2274

Strengths and weaknesses

Categories where Granite 4.0 H Small places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

Granite 4.0 H Small: strongest categories
CategoryScorevs medianRank
Reasoning24.4+0.8#163 of 350, top 47%
Instruction Following68.7−2.5#183 of 305, top 60%
Coding36.4−2.3#209 of 340, top 62%

Weakest categories

Granite 4.0 H Small: weakest categories
CategoryScorevs medianRank
Multilingual38.6−8.8#228 of 297, top 77%
Writing & Preference43.8−10.0#227 of 312, top 73%
Long Context37.7−3.3#213 of 296, top 72%

Closest competitors

The models ranked just above and below Granite 4.0 H Small. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Granite 4.0 H Small
ModelRankScoreBlended $/MSpeed
Qwen Plus#21037.1$0.6037Compare
Gemini 2.5 Flash-Lite#21137.0$0.18172Compare
o3-mini#21236.7$1.93—Compare
Llama 3.1 Nemotron Ultra 253b v1#21336.7——Compare
Command A#21536.5$4.3828Compare
Grok Build 0.1#21636.4$1.25—Compare
gpt-oss-120b#21736.3$0.070355Compare
Mistral Medium#21836.3$368Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

Granite 4.0 H Small Coding benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Coding1249#226 of 294, top 77%LMArena2026-10-08

Reasoning

Granite 4.0 H Small Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Hard Prompts1240#227 of 297, top 77%LMArena2026-10-08

Math

Granite 4.0 H Small Math benchmark results
BenchmarkScorePositionSettingSourceDate
Omni-MATH29.6%#38 of 57, top 67%HELM Capabilities
LMArena Math1247#215 of 285, top 76%LMArena2026-10-08

Knowledge

Granite 4.0 H Small Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
MMLU-Pro56.9%#47 of 58, top 82%HELM Capabilities
Vectara Hallucination Rate (lower is better)5.2%#12 of 96, top 13%Vectara Hallucination Leaderboard
GPQA (HELM)38.3%#46 of 57, top 81%HELM Capabilities
LMArena Expert1251#195 of 273, top 72%LMArena2026-10-08

Multilingual

Granite 4.0 H Small Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1216#228 of 297, top 77%LMArena2026-10-08
LMArena Chinese1249#207 of 285, top 73%LMArena2026-10-08
LMArena Russian1201#233 of 283, top 83%LMArena2026-10-08
LMArena Spanish1258#180 of 226, top 80%LMArena2026-10-08

Instruction Following

Granite 4.0 H Small Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
IFEval89%#11 of 57, top 20%HELM Capabilities
LMArena Instruction Following1222#229 of 298, top 77%LMArena2026-10-08

Long Context

Granite 4.0 H Small Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Longer Query1242#225 of 291, top 78%LMArena2026-10-08

Writing & Preference

Granite 4.0 H Small Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1241#227 of 297, top 77%LMArena2026-10-08
LMArena Creative Writing1211#227 of 295, top 77%LMArena2026-10-08
WildBench73.9%#47 of 57, top 83%HELM Capabilities
LMArena Multi-Turn1242#226 of 295, top 77%LMArena2026-10-08

Compare Granite 4.0 H Small

Other IBM models

Frequently asked questions

How good is Granite 4.0 H Small?

Granite 4.0 H Small by IBM ranks 214th of 354 ranked models on the Noometry Index as of October 2026, with a score of 36.5. Its strongest category is reasoning, where it ranks 163rd.

Is Granite 4.0 H Small open source?

Yes. Granite 4.0 H Small's weights are downloadable; check the license for commercial terms.

What are Granite 4.0 H Small's strengths and weaknesses?

Relative to other ranked models, Granite 4.0 H Small places best in reasoning, instruction following, coding and lowest in multilingual, writing & preference, long context.

What is Granite 4.0 H Small best at?

Its best category is reasoning, where it ranks 163rd on Noometry.