Model comparison
Granite 4.0 Micro vs Llama 3.2 90B
Granite 4.0 Micro is the stronger model overall, scoring 29.0 to 27.5 on the Noometry Index.
Last verified . 2 shared benchmarks.
Summary
- They share 2 benchmarks with published results for both. Granite 4.0 Micro scores higher in 1 category and Llama 3.2 90B in 2 categories; 2 gaps are clear of the uncertainty.
- The widest gap is in knowledge, where Llama 3.2 90B leads 21.7 to 9.9.
- The biggest single-benchmark swing is GPQA Diamond: 28.3% for Granite 4.0 Micro and 41% for Llama 3.2 90B.
Side by side
| Granite 4.0 Micro | Llama 3.2 90B | |
|---|---|---|
| Provider | IBM | Meta |
| Noometry Index | 29.0 | 27.5 |
| Released | 2025-10-02 | 2024-09-24 |
| Weights | Open | Open |
| Context window | 131K | — |
| Max output | 118K | — |
| Input $ / M tokens | $0.017 | — |
| Output $ / M tokens | $0.11 | — |
| Results tracked | 8 | 9 |
Sponsored placements are available on pages like this one. Advertise on Noometry
Category by category
Agentic & Tool Use Not comparable
Granite 4.0 Micro: —, Llama 3.2 90B: 30.0 (#80)
| Benchmark | Granite 4.0 Micro | Llama 3.2 90B |
|---|---|---|
| BALROG | — | 27.3% |
Reasoning Llama 3.2 90B leads
Granite 4.0 Micro: 19.2 (#265), Llama 3.2 90B: 21.7 (#217)
| Benchmark | Granite 4.0 Micro | Llama 3.2 90B |
|---|---|---|
| Chess Puzzles | 0% | — |
| EnigmaEval | — | 0.4% |
| Epoch Capabilities Index | — | 125.5 |
Math Too close to call
Granite 4.0 Micro: 12.0 (#307), Llama 3.2 90B: 11.1 (#308)
| Benchmark | Granite 4.0 Micro | Llama 3.2 90B |
|---|---|---|
| OTIS Mock AIME 2024-2025 | 2.8% | 2.6% |
| Omni-MATH | 20.9% | — |
| MATH Level 5 | — | 39.4% |
Knowledge Llama 3.2 90B leads
Granite 4.0 Micro: 9.9 (#304), Llama 3.2 90B: 21.7 (#274)
| Benchmark | Granite 4.0 Micro | Llama 3.2 90B |
|---|---|---|
| GPQA Diamond | 28.3% | 41% |
| MMLU-Pro | 39.5% | — |
| GPQA (HELM) | 30.7% | — |
| MMLU | — | 80.3% |
Multimodal Not comparable
Granite 4.0 Micro: —, Llama 3.2 90B: 25.4 (#124)
| Benchmark | Granite 4.0 Micro | Llama 3.2 90B |
|---|---|---|
| LMArena Vision | — | 1000 |
| GeoBench | — | 52% |
Instruction Following Not comparable
Granite 4.0 Micro: 69.9 (#169), Llama 3.2 90B: —
| Benchmark | Granite 4.0 Micro | Llama 3.2 90B |
|---|---|---|
| IFEval | 84.9% | — |
Writing & Preference Not comparable
Granite 4.0 Micro: 46.7 (#216), Llama 3.2 90B: —
| Benchmark | Granite 4.0 Micro | Llama 3.2 90B |
|---|---|---|
| WildBench | 67% | — |
Frequently asked questions
Is Granite 4.0 Micro better than Llama 3.2 90B?
Granite 4.0 Micro is the stronger model overall, scoring 29.0 to 27.5 on the Noometry Index.
How many benchmarks do Granite 4.0 Micro and Llama 3.2 90B share?
2 benchmarks have published results for both models. Granite 4.0 Micro has 8 scored results on Noometry and Llama 3.2 90B has 9.