Model comparison

Granite 4.1 8b vs Grok 2 Mini 2024 08 13

Granite 4.1 8b and Grok 2 Mini 2024 08 13 score almost the same on the Noometry Index (37.4 vs 37.7), so choose on price, context window or the category you care about most.

Last verified . 12 shared benchmarks.

Granite 4.1 8b IBM

37.4

Rank #205 Confirmed

Grok 2 Mini 2024 08 13 xAI

37.7

Rank #198 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Granite 4.1 8b scores higher in 7 categories and Grok 2 Mini 2024 08 13 in 1 category; 4 gaps are clear of the uncertainty.
  • The widest gap is in coding, where Grok 2 Mini 2024 08 13 leads 37.0 to 30.1.
  • Granite 4.1 8b has downloadable open weights; the other is API-only.

Side by side

Granite 4.1 8b and Grok 2 Mini 2024 08 13 specifications
Granite 4.1 8bGrok 2 Mini 2024 08 13
ProviderIBMxAI
Noometry Index37.437.7
Released—2024-08-13
WeightsOpenProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked1317

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Grok 2 Mini 2024 08 13 leads

Granite 4.1 8b: 30.1 (#297), Grok 2 Mini 2024 08 13: 37.0 (#199)

Coding benchmarks
BenchmarkGranite 4.1 8bGrok 2 Mini 2024 08 13
LMArena Coding13121269
LMArena WebDev1192—

Reasoning Too close to call

Granite 4.1 8b: 25.7 (#143), Grok 2 Mini 2024 08 13: 24.8 (#159)

Reasoning benchmarks
BenchmarkGranite 4.1 8bGrok 2 Mini 2024 08 13
LMArena Hard Prompts12931255

Math Granite 4.1 8b leads

Granite 4.1 8b: 36.4 (#166), Grok 2 Mini 2024 08 13: 35.4 (#185)

Math benchmarks
BenchmarkGranite 4.1 8bGrok 2 Mini 2024 08 13
LMArena Math13121265

Knowledge Granite 4.1 8b leads

Granite 4.1 8b: 36.1 (#174), Grok 2 Mini 2024 08 13: 33.9 (#200)

Knowledge benchmarks
BenchmarkGranite 4.1 8bGrok 2 Mini 2024 08 13
LMArena Expert13091238

Multilingual Too close to call

Granite 4.1 8b: 41.7 (#204), Grok 2 Mini 2024 08 13: 41.4 (#208)

Multilingual benchmarks
BenchmarkGranite 4.1 8bGrok 2 Mini 2024 08 13
LMArena Non-English12611257
LMArena Chinese13371262
LMArena Russian12401261
LMArena French—1286
LMArena German—1273
LMArena Japanese—1213
LMArena Korean—1195
LMArena Spanish—1279

Instruction Following Granite 4.1 8b leads

Granite 4.1 8b: 66.9 (#203), Grok 2 Mini 2024 08 13: 65.5 (#220)

Instruction Following benchmarks
BenchmarkGranite 4.1 8bGrok 2 Mini 2024 08 13
LMArena Instruction Following12691245

Long Context Too close to call

Granite 4.1 8b: 38.7 (#193), Grok 2 Mini 2024 08 13: 38.4 (#196)

Long Context benchmarks
BenchmarkGranite 4.1 8bGrok 2 Mini 2024 08 13
LMArena Longer Query12751266

Writing & Preference Too close to call

Granite 4.1 8b: 48.1 (#204), Grok 2 Mini 2024 08 13: 47.3 (#211)

Writing & Preference benchmarks
BenchmarkGranite 4.1 8bGrok 2 Mini 2024 08 13
LMArena Text12901281
LMArena Creative Writing12511242
LMArena Multi-Turn12681265

Frequently asked questions

Is Granite 4.1 8b better than Grok 2 Mini 2024 08 13?

Granite 4.1 8b and Grok 2 Mini 2024 08 13 score almost the same on the Noometry Index (37.4 vs 37.7), so choose on price, context window or the category you care about most.

Is Granite 4.1 8b or Grok 2 Mini 2024 08 13 better for coding?

Grok 2 Mini 2024 08 13 scores higher on coding benchmarks: 37.0 versus 30.1 in the Noometry coding category.

How many benchmarks do Granite 4.1 8b and Grok 2 Mini 2024 08 13 share?

12 benchmarks have published results for both models. Granite 4.1 8b has 13 scored results on Noometry and Grok 2 Mini 2024 08 13 has 17.

Related comparisons

Go deeper