Model comparison

Granite 4.0 Micro vs Llama 3.2 90B

Granite 4.0 Micro is the stronger model overall, scoring 29.0 to 27.5 on the Noometry Index.

Last verified . 2 shared benchmarks.

Granite 4.0 Micro IBM

29.0

Rank #318 Confirmed

Llama 3.2 90B Meta

27.5

Rank #331 Confirmed

Summary

  • They share 2 benchmarks with published results for both. Granite 4.0 Micro scores higher in 1 category and Llama 3.2 90B in 2 categories; 2 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Llama 3.2 90B leads 21.7 to 9.9.
  • The biggest single-benchmark swing is GPQA Diamond: 28.3% for Granite 4.0 Micro and 41% for Llama 3.2 90B.

Side by side

Granite 4.0 Micro and Llama 3.2 90B specifications
Granite 4.0 MicroLlama 3.2 90B
ProviderIBMMeta
Noometry Index29.027.5
Released2025-10-022024-09-24
WeightsOpenOpen
Context window131K—
Max output118K—
Input $ / M tokens$0.017—
Output $ / M tokens$0.11—
Results tracked89

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Agentic & Tool Use Not comparable

Granite 4.0 Micro: —, Llama 3.2 90B: 30.0 (#80)

Agentic & Tool Use benchmarks
BenchmarkGranite 4.0 MicroLlama 3.2 90B
BALROG—27.3%

Reasoning Llama 3.2 90B leads

Granite 4.0 Micro: 19.2 (#265), Llama 3.2 90B: 21.7 (#217)

Reasoning benchmarks
BenchmarkGranite 4.0 MicroLlama 3.2 90B
Chess Puzzles0%—
EnigmaEval—0.4%
Epoch Capabilities Index—125.5

Math Too close to call

Granite 4.0 Micro: 12.0 (#307), Llama 3.2 90B: 11.1 (#308)

Math benchmarks
BenchmarkGranite 4.0 MicroLlama 3.2 90B
OTIS Mock AIME 2024-20252.8%2.6%
Omni-MATH20.9%—
MATH Level 5—39.4%

Knowledge Llama 3.2 90B leads

Granite 4.0 Micro: 9.9 (#304), Llama 3.2 90B: 21.7 (#274)

Knowledge benchmarks
BenchmarkGranite 4.0 MicroLlama 3.2 90B
GPQA Diamond28.3%41%
MMLU-Pro39.5%—
GPQA (HELM)30.7%—
MMLU—80.3%

Multimodal Not comparable

Granite 4.0 Micro: —, Llama 3.2 90B: 25.4 (#124)

Multimodal benchmarks
BenchmarkGranite 4.0 MicroLlama 3.2 90B
LMArena Vision—1000
GeoBench—52%

Instruction Following Not comparable

Granite 4.0 Micro: 69.9 (#169), Llama 3.2 90B: —

Instruction Following benchmarks
BenchmarkGranite 4.0 MicroLlama 3.2 90B
IFEval84.9%—

Writing & Preference Not comparable

Granite 4.0 Micro: 46.7 (#216), Llama 3.2 90B: —

Writing & Preference benchmarks
BenchmarkGranite 4.0 MicroLlama 3.2 90B
WildBench67%—

Frequently asked questions

Is Granite 4.0 Micro better than Llama 3.2 90B?

Granite 4.0 Micro is the stronger model overall, scoring 29.0 to 27.5 on the Noometry Index.

How many benchmarks do Granite 4.0 Micro and Llama 3.2 90B share?

2 benchmarks have published results for both models. Granite 4.0 Micro has 8 scored results on Noometry and Llama 3.2 90B has 9.

Related comparisons

Go deeper