Model comparison

Granite 4.0 Micro vs Mercury

Mercury is the stronger model overall, scoring 37.6 to 29.0 on the Noometry Index.

Last verified . 0 shared benchmarks.

Granite 4.0 Micro IBM

29.0

Rank #318 Confirmed

Mercury Inception

37.6

Rank #199 Confirmed

Summary

  • The widest gap is in instruction following, where Granite 4.0 Micro leads 69.9 to 65.2.
  • Granite 4.0 Micro has downloadable open weights; the other is API-only.

Side by side

Granite 4.0 Micro and Mercury specifications
Granite 4.0 MicroMercury
ProviderIBMInception
Noometry Index29.037.6
Released2025-10-02—
WeightsOpenProprietary
Context window131K—
Max output118K—
Input $ / M tokens$0.017—
Output $ / M tokens$0.11—
Results tracked89

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Granite 4.0 Micro: —, Mercury: 38.7 (#170)

Coding benchmarks
BenchmarkGranite 4.0 MicroMercury
LMArena Coding—1322

Reasoning Granite 4.0 Micro leads

Granite 4.0 Micro: 19.2 (#265), Mercury: 17.5 (#293)

Reasoning benchmarks
BenchmarkGranite 4.0 MicroMercury
Kagi LLM Benchmark—21.6%
Chess Puzzles0%—
LMArena Hard Prompts—1285

Math Not comparable

Granite 4.0 Micro: 12.0 (#307), Mercury: —

Math benchmarks
BenchmarkGranite 4.0 MicroMercury
OTIS Mock AIME 2024-20252.8%—
Omni-MATH20.9%—

Knowledge Not comparable

Granite 4.0 Micro: 9.9 (#304), Mercury: —

Knowledge benchmarks
BenchmarkGranite 4.0 MicroMercury
GPQA Diamond28.3%—
MMLU-Pro39.5%—
GPQA (HELM)30.7%—

Multilingual Not comparable

Granite 4.0 Micro: —, Mercury: 41.6 (#206)

Multilingual benchmarks
BenchmarkGranite 4.0 MicroMercury
LMArena Non-English—1260

Instruction Following Granite 4.0 Micro leads

Granite 4.0 Micro: 69.9 (#169), Mercury: 65.2 (#224)

Instruction Following benchmarks
BenchmarkGranite 4.0 MicroMercury
IFEval84.9%—
LMArena Instruction Following—1239

Long Context Not comparable

Granite 4.0 Micro: —, Mercury: 38.4 (#198)

Long Context benchmarks
BenchmarkGranite 4.0 MicroMercury
LMArena Longer Query—1266

Writing & Preference Too close to call

Granite 4.0 Micro: 46.7 (#216), Mercury: 46.2 (#221)

Writing & Preference benchmarks
BenchmarkGranite 4.0 MicroMercury
LMArena Text—1282
LMArena Creative Writing—1191
WildBench67%—
LMArena Multi-Turn—1282

Frequently asked questions

Is Granite 4.0 Micro better than Mercury?

Mercury is the stronger model overall, scoring 37.6 to 29.0 on the Noometry Index.

How many benchmarks do Granite 4.0 Micro and Mercury share?

0 benchmarks have published results for both models. Granite 4.0 Micro has 8 scored results on Noometry and Mercury has 9.

Related comparisons

Go deeper