Model comparison

Amazon Nova Pro vs Granite 4.2 3b

Granite 4.2 3b is the stronger model overall, scoring 39.4 to 31.0 on the Noometry Index.

Last verified . 11 shared benchmarks.

Amazon Nova Pro Amazon

31.0

Rank #281 Confirmed

Granite 4.2 3b IBM

39.4

Rank #169 Confirmed

Summary

  • They share 11 benchmarks with published results for both. Amazon Nova Pro scores higher in 0 categories and Granite 4.2 3b in 7 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Granite 4.2 3b leads 36.3 to 27.4.
  • Granite 4.2 3b has downloadable open weights; the other is API-only.

Side by side

Amazon Nova Pro and Granite 4.2 3b specifications
Amazon Nova ProGranite 4.2 3b
ProviderAmazonIBM
Noometry Index31.039.4
Released2024-12-03—
WeightsProprietaryOpen
Context window300K—
Max output10K—
Input $ / M tokens$0.80—
Output $ / M tokens$3.20—
Results tracked3811

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Granite 4.2 3b leads

Amazon Nova Pro: 35.1 (#229), Granite 4.2 3b: 40.0 (#151)

Coding benchmarks
BenchmarkAmazon Nova ProGranite 4.2 3b
LMArena Coding12701361
LiveBench Coding38.1%—

Agentic & Tool Use Not comparable

Amazon Nova Pro: 16.7 (#147), Granite 4.2 3b: —

Agentic & Tool Use benchmarks
BenchmarkAmazon Nova ProGranite 4.2 3b
Berkeley Function Calling Leaderboard25%—
TheAgentCompany1.7%—

Reasoning Granite 4.2 3b leads

Amazon Nova Pro: 20.0 (#243), Granite 4.2 3b: 26.0 (#138)

Reasoning benchmarks
BenchmarkAmazon Nova ProGranite 4.2 3b
LMArena Hard Prompts12461306
LiveBench Reasoning32.6%—
LiveBench Data Analysis48.3%—
Epoch Capabilities Index123.8—
LiveBench43.5%—

Math Not comparable

Amazon Nova Pro: 28.5 (#243), Granite 4.2 3b: —

Math benchmarks
BenchmarkAmazon Nova ProGranite 4.2 3b
Omni-MATH24.2%—
LiveBench Math38%—
LMArena Math1252—

Knowledge Granite 4.2 3b leads

Amazon Nova Pro: 27.4 (#250), Granite 4.2 3b: 36.3 (#171)

Knowledge benchmarks
BenchmarkAmazon Nova ProGranite 4.2 3b
LMArena Expert12111315
Humanity's Last Exam4.4%—
MMLU-Pro67.3%—
Confabulations30.1%—
Vectara Hallucination Rate5.1%—
GPQA (HELM)44.6%—
MMLU82%—

Multimodal Not comparable

Amazon Nova Pro: 25.0 (#126), Granite 4.2 3b: —

Multimodal benchmarks
BenchmarkAmazon Nova ProGranite 4.2 3b
LMArena Vision980—

Multilingual Granite 4.2 3b leads

Amazon Nova Pro: 39.7 (#223), Granite 4.2 3b: 42.1 (#198)

Multilingual benchmarks
BenchmarkAmazon Nova ProGranite 4.2 3b
LMArena Non-English12341268
LMArena Chinese12441269
LMArena Russian12401249
LMArena French1271—
LMArena German1243—
LMArena Japanese1200—
LMArena Korean1203—
LMArena Spanish1182—

Instruction Following Granite 4.2 3b leads

Amazon Nova Pro: 64.9 (#226), Granite 4.2 3b: 67.1 (#200)

Instruction Following benchmarks
BenchmarkAmazon Nova ProGranite 4.2 3b
LMArena Instruction Following12351273
LiveBench Instruction Following67.1%—
IFEval81.5%—

Long Context Granite 4.2 3b leads

Amazon Nova Pro: 38.1 (#205), Granite 4.2 3b: 39.2 (#185)

Long Context benchmarks
BenchmarkAmazon Nova ProGranite 4.2 3b
LMArena Longer Query12551291

Writing & Preference Granite 4.2 3b leads

Amazon Nova Pro: 43.9 (#226), Granite 4.2 3b: 47.2 (#212)

Writing & Preference benchmarks
BenchmarkAmazon Nova ProGranite 4.2 3b
LMArena Text12591293
LMArena Creative Writing12121205
LMArena Multi-Turn12461290
Short-Story Creative Writing60.5%—
WildBench77.7%—
LiveBench Language37%—

Frequently asked questions

Is Amazon Nova Pro better than Granite 4.2 3b?

Granite 4.2 3b is the stronger model overall, scoring 39.4 to 31.0 on the Noometry Index.

Is Amazon Nova Pro or Granite 4.2 3b better for coding?

Granite 4.2 3b scores higher on coding benchmarks: 40.0 versus 35.1 in the Noometry coding category.

How many benchmarks do Amazon Nova Pro and Granite 4.2 3b share?

11 benchmarks have published results for both models. Amazon Nova Pro has 38 scored results on Noometry and Granite 4.2 3b has 11.

Related comparisons

Go deeper