Model comparison

Granite 3.0 2b Instruct vs Granite 4.2 30b

Granite 4.2 30b is the stronger model overall, scoring 41.8 to 30.8 on the Noometry Index.

Last verified . 11 shared benchmarks.

Granite 3.0 2b Instruct IBM

30.8

Rank #286 Confirmed

Granite 4.2 30b IBM

41.8

Rank #130 Confirmed

Summary

  • They share 11 benchmarks with published results for both. Granite 3.0 2b Instruct scores higher in 0 categories and Granite 4.2 30b in 7 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Granite 4.2 30b leads 53.8 to 29.6.

Side by side

Granite 3.0 2b Instruct and Granite 4.2 30b specifications
Granite 3.0 2b InstructGranite 4.2 30b
ProviderIBMIBM
Noometry Index30.841.8
Released——
WeightsOpenOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked1311

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Granite 4.2 30b leads

Granite 3.0 2b Instruct: 28.3 (#316), Granite 4.2 30b: 41.0 (#126)

Coding benchmarks
BenchmarkGranite 3.0 2b InstructGranite 4.2 30b
LMArena Coding10901396
BigCodeBench Instruct20.5%—

Reasoning Granite 4.2 30b leads

Granite 3.0 2b Instruct: 20.5 (#235), Granite 4.2 30b: 27.8 (#112)

Reasoning benchmarks
BenchmarkGranite 3.0 2b InstructGranite 4.2 30b
LMArena Hard Prompts10731374

Math Not comparable

Granite 3.0 2b Instruct: 32.2 (#217), Granite 4.2 30b: —

Math benchmarks
BenchmarkGranite 3.0 2b InstructGranite 4.2 30b
LMArena Math1117—

Knowledge Granite 4.2 30b leads

Granite 3.0 2b Instruct: 29.0 (#241), Granite 4.2 30b: 39.1 (#138)

Knowledge benchmarks
BenchmarkGranite 3.0 2b InstructGranite 4.2 30b
LMArena Expert10641406

Multilingual Granite 4.2 30b leads

Granite 3.0 2b Instruct: 27.0 (#278), Granite 4.2 30b: 47.3 (#151)

Multilingual benchmarks
BenchmarkGranite 3.0 2b InstructGranite 4.2 30b
LMArena Non-English10331340
LMArena Chinese10701414
LMArena Russian10451343

Instruction Following Granite 4.2 30b leads

Granite 3.0 2b Instruct: 53.9 (#284), Granite 4.2 30b: 71.2 (#155)

Instruction Following benchmarks
BenchmarkGranite 3.0 2b InstructGranite 4.2 30b
LMArena Instruction Following10561347

Long Context Granite 4.2 30b leads

Granite 3.0 2b Instruct: 32.5 (#268), Granite 4.2 30b: 41.4 (#140)

Long Context benchmarks
BenchmarkGranite 3.0 2b InstructGranite 4.2 30b
LMArena Longer Query10701359

Writing & Preference Granite 4.2 30b leads

Granite 3.0 2b Instruct: 29.6 (#292), Granite 4.2 30b: 53.8 (#156)

Writing & Preference benchmarks
BenchmarkGranite 3.0 2b InstructGranite 4.2 30b
LMArena Text10801361
LMArena Creative Writing10461288
LMArena Multi-Turn10531339

Frequently asked questions

Is Granite 3.0 2b Instruct better than Granite 4.2 30b?

Granite 4.2 30b is the stronger model overall, scoring 41.8 to 30.8 on the Noometry Index.

Is Granite 3.0 2b Instruct or Granite 4.2 30b better for coding?

Granite 4.2 30b scores higher on coding benchmarks: 41.0 versus 28.3 in the Noometry coding category.

How many benchmarks do Granite 3.0 2b Instruct and Granite 4.2 30b share?

11 benchmarks have published results for both models. Granite 3.0 2b Instruct has 13 scored results on Noometry and Granite 4.2 30b has 11.

Related comparisons

Go deeper