IBM, open weights
Granite 4.0 H Small
Granite 4.0 H Small by IBM ranks 214th of 354 ranked models on the Noometry Index as of October 2026, with a score of 36.5. Its strongest category is reasoning, where it ranks 163rd.
Last verified
Specifications
- Noometry rank
- #214 of 354
- Index score
- 36.5
- Evidence
- Confirmed 19 results
- Provider
- IBM
- Released
- Unknown
- Weights
- Open weights
- Reasoning
- Unknown
- Context window
- —
- Max output
- —
- Input price
- Not listed
- Output price
- Not listed
- Blended price
- Not listed
- Output speed
- Not measured
- Value
- Not ranked
- Knowledge cutoff
- Unknown
Category scores
Each category score combines every public result we have in that category.
- Coding 36.4
- Reasoning 24.4
- Math 31.4
- Knowledge 32.0
- Multilingual 38.6
- Instruction Following 68.7
- Long Context 37.7
- Writing & Preference 43.8
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 36.4 | #209 | 1 |
| Reasoning | 24.4 | #163 | 1 |
| Math | 31.4 | #223 | 2 |
| Knowledge | 32.0 | #215 | 4 |
| Multilingual | 38.6 | #228 | 1 |
| Instruction Following | 68.7 | #183 | 2 |
| Long Context | 37.7 | #213 | 1 |
| Writing & Preference | 43.8 | #227 | 4 |
Strengths and weaknesses
Categories where Granite 4.0 H Small places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Reasoning | 24.4 | +0.8 | #163 of 350, top 47% |
| Instruction Following | 68.7 | −2.5 | #183 of 305, top 60% |
| Coding | 36.4 | −2.3 | #209 of 340, top 62% |
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Multilingual | 38.6 | −8.8 | #228 of 297, top 77% |
| Writing & Preference | 43.8 | −10.0 | #227 of 312, top 73% |
| Long Context | 37.7 | −3.3 | #213 of 296, top 72% |
Closest competitors
The models ranked just above and below Granite 4.0 H Small. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Qwen Plus | #210 | 37.1 | $0.60 | 37 | Compare |
| Gemini 2.5 Flash-Lite | #211 | 37.0 | $0.18 | 172 | Compare |
| o3-mini | #212 | 36.7 | $1.93 | — | Compare |
| Llama 3.1 Nemotron Ultra 253b v1 | #213 | 36.7 | — | — | Compare |
| Command A | #215 | 36.5 | $4.38 | 28 | Compare |
| Grok Build 0.1 | #216 | 36.4 | $1.25 | — | Compare |
| gpt-oss-120b | #217 | 36.3 | $0.0703 | 55 | Compare |
| Mistral Medium | #218 | 36.3 | $3 | 68 | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Coding | 1249 | #226 of 294, top 77% | LMArena | 2026-10-08 |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Hard Prompts | 1240 | #227 of 297, top 77% | LMArena | 2026-10-08 |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| Omni-MATH | 29.6% | #38 of 57, top 67% | HELM Capabilities | ||
| LMArena Math | 1247 | #215 of 285, top 76% | LMArena | 2026-10-08 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| MMLU-Pro | 56.9% | #47 of 58, top 82% | HELM Capabilities | ||
| Vectara Hallucination Rate (lower is better) | 5.2% | #12 of 96, top 13% | Vectara Hallucination Leaderboard | ||
| GPQA (HELM) | 38.3% | #46 of 57, top 81% | HELM Capabilities | ||
| LMArena Expert | 1251 | #195 of 273, top 72% | LMArena | 2026-10-08 |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1216 | #228 of 297, top 77% | LMArena | 2026-10-08 | |
| LMArena Chinese | 1249 | #207 of 285, top 73% | LMArena | 2026-10-08 | |
| LMArena Russian | 1201 | #233 of 283, top 83% | LMArena | 2026-10-08 | |
| LMArena Spanish | 1258 | #180 of 226, top 80% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| IFEval | 89% | #11 of 57, top 20% | HELM Capabilities | ||
| LMArena Instruction Following | 1222 | #229 of 298, top 77% | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 1242 | #225 of 291, top 78% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1241 | #227 of 297, top 77% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1211 | #227 of 295, top 77% | LMArena | 2026-10-08 | |
| WildBench | 73.9% | #47 of 57, top 83% | HELM Capabilities | ||
| LMArena Multi-Turn | 1242 | #226 of 295, top 77% | LMArena | 2026-10-08 |
Compare Granite 4.0 H Small
- Granite 4.0 H Small vs Llama 3.1 Nemotron Ultra 253b v1
- Granite 4.0 H Small vs Command A
- Granite 4.0 H Small vs o3-mini
- Granite 4.0 H Small vs Grok Build 0.1
- Granite 4.0 H Small vs Gemini 2.5 Flash-Lite
- Granite 4.0 H Small vs gpt-oss-120b
- Granite 4.0 H Small vs GPT-6 Astra
- Granite 4.0 H Small vs Claude Fable 5.1
- Granite 4.0 H Small vs Gemini 3.8 Flash
- Granite 4.0 H Small vs Kimi K3
- Granite 4.0 H Small vs Grok 4.6
- Granite 4.0 H Small vs Qwen3.8 Max
- Granite 4.0 H Small vs GLM-5.3
- Granite 4.0 H Small vs Muse Spark 1.3
Other IBM models
Frequently asked questions
How good is Granite 4.0 H Small?
Granite 4.0 H Small by IBM ranks 214th of 354 ranked models on the Noometry Index as of October 2026, with a score of 36.5. Its strongest category is reasoning, where it ranks 163rd.
Is Granite 4.0 H Small open source?
Yes. Granite 4.0 H Small's weights are downloadable; check the license for commercial terms.
What are Granite 4.0 H Small's strengths and weaknesses?
Relative to other ranked models, Granite 4.0 H Small places best in reasoning, instruction following, coding and lowest in multilingual, writing & preference, long context.
What is Granite 4.0 H Small best at?
Its best category is reasoning, where it ranks 163rd on Noometry.