Model comparison

Codellama 70b Instruct vs Granite 4.2 8B

Granite 4.2 8B is the stronger model overall, scoring 40.5 to 33.7 on the Noometry Index.

Last verified . 4 shared benchmarks.

Codellama 70b Instruct Meta

33.7

Rank #237 Confirmed

Granite 4.2 8B IBM

40.5

Rank #148 Confirmed

Summary

  • They share 4 benchmarks with published results for both. Codellama 70b Instruct scores higher in 0 categories and Granite 4.2 8B in 5 categories; 5 gaps are clear of the uncertainty.
  • The widest gap is in multilingual, where Granite 4.2 8B leads 44.5 to 24.8.

Side by side

Codellama 70b Instruct and Granite 4.2 8B specifications
Codellama 70b InstructGranite 4.2 8B
ProviderMetaIBM
Noometry Index33.740.5
Released——
WeightsOpenOpen
Context window—131K
Max output—118K
Input $ / M tokens—$0.06
Output $ / M tokens—$0.25
Results tracked711

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Granite 4.2 8B leads

Codellama 70b Instruct: 37.6 (#193), Granite 4.2 8B: 40.5 (#137)

Coding benchmarks
BenchmarkCodellama 70b InstructGranite 4.2 8B
BigCodeBench Instruct40.7%—
LMArena Coding—1380
BigCodeBench Complete49.6%—
HumanEval+65.9%—

Reasoning Granite 4.2 8B leads

Codellama 70b Instruct: 20.1 (#242), Granite 4.2 8B: 26.6 (#131)

Reasoning benchmarks
BenchmarkCodellama 70b InstructGranite 4.2 8B
LMArena Hard Prompts10521329

Knowledge Not comparable

Codellama 70b Instruct: —, Granite 4.2 8B: 38.4 (#145)

Knowledge benchmarks
BenchmarkCodellama 70b InstructGranite 4.2 8B
LMArena Expert—1384

Multilingual Granite 4.2 8B leads

Codellama 70b Instruct: 24.8 (#288), Granite 4.2 8B: 44.5 (#178)

Multilingual benchmarks
BenchmarkCodellama 70b InstructGranite 4.2 8B
LMArena Non-English9921302
LMArena Chinese—1366
LMArena Russian—1285

Instruction Following Granite 4.2 8B leads

Codellama 70b Instruct: 51.9 (#293), Granite 4.2 8B: 68.7 (#184)

Instruction Following benchmarks
BenchmarkCodellama 70b InstructGranite 4.2 8B
LMArena Instruction Following10241301

Long Context Not comparable

Codellama 70b Instruct: —, Granite 4.2 8B: 40.3 (#159)

Long Context benchmarks
BenchmarkCodellama 70b InstructGranite 4.2 8B
LMArena Longer Query—1324

Writing & Preference Granite 4.2 8B leads

Codellama 70b Instruct: 33.4 (#277), Granite 4.2 8B: 49.6 (#189)

Writing & Preference benchmarks
BenchmarkCodellama 70b InstructGranite 4.2 8B
LMArena Text10571320
LMArena Creative Writing—1236
LMArena Multi-Turn—1301

Frequently asked questions

Is Codellama 70b Instruct better than Granite 4.2 8B?

Granite 4.2 8B is the stronger model overall, scoring 40.5 to 33.7 on the Noometry Index.

Is Codellama 70b Instruct or Granite 4.2 8B better for coding?

Granite 4.2 8B scores higher on coding benchmarks: 40.5 versus 37.6 in the Noometry coding category.

How many benchmarks do Codellama 70b Instruct and Granite 4.2 8B share?

4 benchmarks have published results for both models. Codellama 70b Instruct has 7 scored results on Noometry and Granite 4.2 8B has 11.

Related comparisons

Go deeper