Model comparison

Granite 4.2 3b vs Mercury

Granite 4.2 3b is the stronger model overall, scoring 39.4 to 37.6 on the Noometry Index.

Last verified . 8 shared benchmarks.

Granite 4.2 3b IBM

39.4

Rank #169 Confirmed

Mercury Inception

37.6

Rank #199 Confirmed

Summary

  • They share 8 benchmarks with published results for both. Granite 4.2 3b scores higher in 6 categories and Mercury in 0 categories; 4 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Granite 4.2 3b leads 26.0 to 17.5.
  • Granite 4.2 3b has downloadable open weights; the other is API-only.

Side by side

Granite 4.2 3b and Mercury specifications
Granite 4.2 3bMercury
ProviderIBMInception
Noometry Index39.437.6
Released——
WeightsOpenProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked119

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Granite 4.2 3b leads

Granite 4.2 3b: 40.0 (#151), Mercury: 38.7 (#170)

Coding benchmarks
BenchmarkGranite 4.2 3bMercury
LMArena Coding13611322

Reasoning Granite 4.2 3b leads

Granite 4.2 3b: 26.0 (#138), Mercury: 17.5 (#293)

Reasoning benchmarks
BenchmarkGranite 4.2 3bMercury
LMArena Hard Prompts13061285
Kagi LLM Benchmark—21.6%

Knowledge Not comparable

Granite 4.2 3b: 36.3 (#171), Mercury: —

Knowledge benchmarks
BenchmarkGranite 4.2 3bMercury
LMArena Expert1315—

Multilingual Too close to call

Granite 4.2 3b: 42.1 (#198), Mercury: 41.6 (#206)

Multilingual benchmarks
BenchmarkGranite 4.2 3bMercury
LMArena Non-English12681260
LMArena Chinese1269—
LMArena Russian1249—

Instruction Following Granite 4.2 3b leads

Granite 4.2 3b: 67.1 (#200), Mercury: 65.2 (#224)

Instruction Following benchmarks
BenchmarkGranite 4.2 3bMercury
LMArena Instruction Following12731239

Long Context Too close to call

Granite 4.2 3b: 39.2 (#185), Mercury: 38.4 (#198)

Long Context benchmarks
BenchmarkGranite 4.2 3bMercury
LMArena Longer Query12911266

Writing & Preference Granite 4.2 3b leads

Granite 4.2 3b: 47.2 (#212), Mercury: 46.2 (#221)

Writing & Preference benchmarks
BenchmarkGranite 4.2 3bMercury
LMArena Text12931282
LMArena Creative Writing12051191
LMArena Multi-Turn12901282

Frequently asked questions

Is Granite 4.2 3b better than Mercury?

Granite 4.2 3b is the stronger model overall, scoring 39.4 to 37.6 on the Noometry Index.

Is Granite 4.2 3b or Mercury better for coding?

Granite 4.2 3b scores higher on coding benchmarks: 40.0 versus 38.7 in the Noometry coding category.

How many benchmarks do Granite 4.2 3b and Mercury share?

8 benchmarks have published results for both models. Granite 4.2 3b has 11 scored results on Noometry and Mercury has 9.

Related comparisons

Go deeper