Model comparison

Granite 4.2 30b vs Mercury

Granite 4.2 30b is the stronger model overall, scoring 41.8 to 37.6 on the Noometry Index.

Last verified . 8 shared benchmarks.

Granite 4.2 30b IBM

41.8

Rank #130 Confirmed

Mercury Inception

37.6

Rank #199 Confirmed

Summary

  • They share 8 benchmarks with published results for both. Granite 4.2 30b scores higher in 6 categories and Mercury in 0 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Granite 4.2 30b leads 27.8 to 17.5.
  • Granite 4.2 30b has downloadable open weights; the other is API-only.

Side by side

Granite 4.2 30b and Mercury specifications
Granite 4.2 30bMercury
ProviderIBMInception
Noometry Index41.837.6
Released——
WeightsOpenProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked119

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Granite 4.2 30b leads

Granite 4.2 30b: 41.0 (#126), Mercury: 38.7 (#170)

Coding benchmarks
BenchmarkGranite 4.2 30bMercury
LMArena Coding13961322

Reasoning Granite 4.2 30b leads

Granite 4.2 30b: 27.8 (#112), Mercury: 17.5 (#293)

Reasoning benchmarks
BenchmarkGranite 4.2 30bMercury
LMArena Hard Prompts13741285
Kagi LLM Benchmark—21.6%

Knowledge Not comparable

Granite 4.2 30b: 39.1 (#138), Mercury: —

Knowledge benchmarks
BenchmarkGranite 4.2 30bMercury
LMArena Expert1406—

Multilingual Granite 4.2 30b leads

Granite 4.2 30b: 47.3 (#151), Mercury: 41.6 (#206)

Multilingual benchmarks
BenchmarkGranite 4.2 30bMercury
LMArena Non-English13401260
LMArena Chinese1414—
LMArena Russian1343—

Instruction Following Granite 4.2 30b leads

Granite 4.2 30b: 71.2 (#155), Mercury: 65.2 (#224)

Instruction Following benchmarks
BenchmarkGranite 4.2 30bMercury
LMArena Instruction Following13471239

Long Context Granite 4.2 30b leads

Granite 4.2 30b: 41.4 (#140), Mercury: 38.4 (#198)

Long Context benchmarks
BenchmarkGranite 4.2 30bMercury
LMArena Longer Query13591266

Writing & Preference Granite 4.2 30b leads

Granite 4.2 30b: 53.8 (#156), Mercury: 46.2 (#221)

Writing & Preference benchmarks
BenchmarkGranite 4.2 30bMercury
LMArena Text13611282
LMArena Creative Writing12881191
LMArena Multi-Turn13391282

Frequently asked questions

Is Granite 4.2 30b better than Mercury?

Granite 4.2 30b is the stronger model overall, scoring 41.8 to 37.6 on the Noometry Index.

Is Granite 4.2 30b or Mercury better for coding?

Granite 4.2 30b scores higher on coding benchmarks: 41.0 versus 38.7 in the Noometry coding category.

How many benchmarks do Granite 4.2 30b and Mercury share?

8 benchmarks have published results for both models. Granite 4.2 30b has 11 scored results on Noometry and Mercury has 9.

Related comparisons

Go deeper