Model comparison

Codestral vs Granite 4.2 30b

Granite 4.2 30b is the stronger model overall, scoring 41.8 to 30.6 on the Noometry Index.

Last verified . 0 shared benchmarks.

Codestral Mistral AI

30.6

Rank #290 Reported

Granite 4.2 30b IBM

41.8

Rank #130 Confirmed

Summary

  • The widest gap is in coding, where Granite 4.2 30b leads 41.0 to 27.3.
  • Granite 4.2 30b has downloadable open weights; the other is API-only.

Side by side

Codestral and Granite 4.2 30b specifications
CodestralGranite 4.2 30b
ProviderMistral AIIBM
Noometry Index30.641.8
Released2024-05-29—
WeightsProprietaryOpen
Context window256K—
Max output8K—
Input $ / M tokens$0.30—
Output $ / M tokens$0.90—
Results tracked711

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Granite 4.2 30b leads

Codestral: 27.3 (#321), Granite 4.2 30b: 41.0 (#126)

Coding benchmarks
BenchmarkCodestralGranite 4.2 30b
Aider Polyglot11.1%—
BigCodeBench Instruct41.8%—
LMArena Coding—1396
BigCodeBench Complete52.5%—
ALE-Bench137.78—
HumanEval+73.8%—
MBPP+61.9%—

Reasoning Granite 4.2 30b leads

Codestral: 19.8 (#251), Granite 4.2 30b: 27.8 (#112)

Reasoning benchmarks
BenchmarkCodestralGranite 4.2 30b
Kagi LLM Benchmark32.5%—
LMArena Hard Prompts—1374

Knowledge Not comparable

Codestral: —, Granite 4.2 30b: 39.1 (#138)

Knowledge benchmarks
BenchmarkCodestralGranite 4.2 30b
LMArena Expert—1406

Multilingual Not comparable

Codestral: —, Granite 4.2 30b: 47.3 (#151)

Multilingual benchmarks
BenchmarkCodestralGranite 4.2 30b
LMArena Non-English—1340
LMArena Chinese—1414
LMArena Russian—1343

Instruction Following Not comparable

Codestral: —, Granite 4.2 30b: 71.2 (#155)

Instruction Following benchmarks
BenchmarkCodestralGranite 4.2 30b
LMArena Instruction Following—1347

Long Context Not comparable

Codestral: —, Granite 4.2 30b: 41.4 (#140)

Long Context benchmarks
BenchmarkCodestralGranite 4.2 30b
LMArena Longer Query—1359

Writing & Preference Not comparable

Codestral: —, Granite 4.2 30b: 53.8 (#156)

Writing & Preference benchmarks
BenchmarkCodestralGranite 4.2 30b
LMArena Text—1361
LMArena Creative Writing—1288
LMArena Multi-Turn—1339

Frequently asked questions

Is Codestral better than Granite 4.2 30b?

Granite 4.2 30b is the stronger model overall, scoring 41.8 to 30.6 on the Noometry Index.

Is Codestral or Granite 4.2 30b better for coding?

Granite 4.2 30b scores higher on coding benchmarks: 41.0 versus 27.3 in the Noometry coding category.

How many benchmarks do Codestral and Granite 4.2 30b share?

0 benchmarks have published results for both models. Codestral has 7 scored results on Noometry and Granite 4.2 30b has 11.

Related comparisons

Go deeper