Model comparison

Granite 4.2 8B vs Inkling

Inkling is the stronger model overall, scoring 44.1 to 40.5 on the Noometry Index. Granite 4.2 8B costs 24× less per token, which makes it the better buy when Inkling's lead doesn't matter for your workload.

Last verified . 11 shared benchmarks.

Granite 4.2 8B IBM

40.5

Rank #148 Confirmed

Inkling Thinking Machines Lab

44.1

Rank #80 Confirmed

Summary

  • They share 11 benchmarks with published results for both. Granite 4.2 8B scores higher in 1 category and Inkling in 6 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Inkling leads 55.1 to 38.4.
  • Granite 4.2 8B is cheaper at $0.06 / $0.25 per million input/output tokens, against $1.87 / $4.68 for Inkling.
  • Granite 4.2 8B accepts more context: 131K tokens versus 66K.

Side by side

Granite 4.2 8B and Inkling specifications
Granite 4.2 8BInkling
ProviderIBMThinking Machines Lab
Noometry Index40.544.1
Released—2026-07-15
WeightsOpenOpen
Context window131K66K
Max output118K66K
Input $ / M tokens$0.06$1.87
Output $ / M tokens$0.25$4.68
Results tracked1141

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Granite 4.2 8B leads

Granite 4.2 8B: 40.5 (#137), Inkling: 34.5 (#234)

Coding benchmarks
BenchmarkGranite 4.2 8BInkling
LMArena Coding13801464
FrontierCode—14%
LMArena WebDev—1413
FrontierSWE—4.1%
SciCode—47%
WeirdML—32.3%
ALE-Bench—946

Agentic & Tool Use Not comparable

Granite 4.2 8B: —, Inkling: 29.6 (#85)

Agentic & Tool Use benchmarks
BenchmarkGranite 4.2 8BInkling
APEX-Agents—33.8%
τ²-bench Banking—25%

Reasoning Inkling leads

Granite 4.2 8B: 26.6 (#131), Inkling: 40.4 (#56)

Reasoning benchmarks
BenchmarkGranite 4.2 8BInkling
LMArena Hard Prompts13291451
ARC-AGI-2—36.5%
SimpleBench—50%
ARC-AGI-1—79.5%
CritPt—5.4%
Chess Puzzles—21%
DTBench—87.5%
LMCA—37.6%
Epoch Capabilities Index—148.54

Math Not comparable

Granite 4.2 8B: —, Inkling: 31.3 (#225)

Math benchmarks
BenchmarkGranite 4.2 8BInkling
FrontierMath (Tiers 1-3)—33.3%
FrontierMath Tier 4—4.9%
OTIS Mock AIME 2024-2025—88.9%
ProofBench—0%
LMArena Math—1479

Knowledge Inkling leads

Granite 4.2 8B: 38.4 (#145), Inkling: 55.1 (#49)

Knowledge benchmarks
BenchmarkGranite 4.2 8BInkling
LMArena Expert13841465
GPQA Diamond—88.3%
SimpleQA Verified—40.3%

Multilingual Inkling leads

Granite 4.2 8B: 44.5 (#178), Inkling: 54.0 (#52)

Multilingual benchmarks
BenchmarkGranite 4.2 8BInkling
LMArena Non-English13021434
LMArena Chinese13661490
LMArena Russian12851429
LMArena French—1458
LMArena German—1446
LMArena Japanese—1429
LMArena Korean—1404
LMArena Spanish—1448

Instruction Following Inkling leads

Granite 4.2 8B: 68.7 (#184), Inkling: 75.1 (#71)

Instruction Following benchmarks
BenchmarkGranite 4.2 8BInkling
LMArena Instruction Following13011426

Long Context Inkling leads

Granite 4.2 8B: 40.3 (#159), Inkling: 43.8 (#86)

Long Context benchmarks
BenchmarkGranite 4.2 8BInkling
LMArena Longer Query13241434

Writing & Preference Inkling leads

Granite 4.2 8B: 49.6 (#189), Inkling: 65.2 (#51)

Writing & Preference benchmarks
BenchmarkGranite 4.2 8BInkling
LMArena Text13201441
LMArena Creative Writing12361387
LMArena Multi-Turn13011436
EQ-Bench Creative Writing—1611
EQ-Bench 4—1226

Frequently asked questions

Is Granite 4.2 8B better than Inkling?

Inkling is the stronger model overall, scoring 44.1 to 40.5 on the Noometry Index. Granite 4.2 8B costs 24× less per token, which makes it the better buy when Inkling's lead doesn't matter for your workload.

Which is cheaper, Granite 4.2 8B or Inkling?

Granite 4.2 8B is cheaper. It lists at $0.06 per million input tokens and $0.25 per million output tokens; Inkling lists at $1.87 and $4.68.

Is Granite 4.2 8B or Inkling better for coding?

Granite 4.2 8B scores higher on coding benchmarks: 40.5 versus 34.5 in the Noometry coding category.

Which has the bigger context window?

Granite 4.2 8B does, with 131K tokens against 66K.

How many benchmarks do Granite 4.2 8B and Inkling share?

11 benchmarks have published results for both models. Granite 4.2 8B has 11 scored results on Noometry and Inkling has 41.

Related comparisons

Go deeper