Model comparison
Codellama 70b Instruct vs Granite 4.0 Micro
Codellama 70b Instruct is the stronger model overall, scoring 33.7 to 29.0 on the Noometry Index.
Last verified . 0 shared benchmarks.
Summary
- The widest gap is in instruction following, where Granite 4.0 Micro leads 69.9 to 51.9.
Side by side
| Codellama 70b Instruct | Granite 4.0 Micro | |
|---|---|---|
| Provider | Meta | IBM |
| Noometry Index | 33.7 | 29.0 |
| Released | — | 2025-10-02 |
| Weights | Open | Open |
| Context window | — | 131K |
| Max output | — | 118K |
| Input $ / M tokens | — | $0.017 |
| Output $ / M tokens | — | $0.11 |
| Results tracked | 7 | 8 |
Sponsored placements are available on pages like this one. Advertise on Noometry
Category by category
Coding Not comparable
Codellama 70b Instruct: 37.6 (#193), Granite 4.0 Micro: —
| Benchmark | Codellama 70b Instruct | Granite 4.0 Micro |
|---|---|---|
| BigCodeBench Instruct | 40.7% | — |
| BigCodeBench Complete | 49.6% | — |
| HumanEval+ | 65.9% | — |
Reasoning Too close to call
Codellama 70b Instruct: 20.1 (#242), Granite 4.0 Micro: 19.2 (#265)
| Benchmark | Codellama 70b Instruct | Granite 4.0 Micro |
|---|---|---|
| Chess Puzzles | — | 0% |
| LMArena Hard Prompts | 1052 | — |
Math Not comparable
Codellama 70b Instruct: —, Granite 4.0 Micro: 12.0 (#307)
| Benchmark | Codellama 70b Instruct | Granite 4.0 Micro |
|---|---|---|
| OTIS Mock AIME 2024-2025 | — | 2.8% |
| Omni-MATH | — | 20.9% |
Knowledge Not comparable
Codellama 70b Instruct: —, Granite 4.0 Micro: 9.9 (#304)
| Benchmark | Codellama 70b Instruct | Granite 4.0 Micro |
|---|---|---|
| GPQA Diamond | — | 28.3% |
| MMLU-Pro | — | 39.5% |
| GPQA (HELM) | — | 30.7% |
Multilingual Not comparable
Codellama 70b Instruct: 24.8 (#288), Granite 4.0 Micro: —
| Benchmark | Codellama 70b Instruct | Granite 4.0 Micro |
|---|---|---|
| LMArena Non-English | 992 | — |
Instruction Following Granite 4.0 Micro leads
Codellama 70b Instruct: 51.9 (#293), Granite 4.0 Micro: 69.9 (#169)
| Benchmark | Codellama 70b Instruct | Granite 4.0 Micro |
|---|---|---|
| IFEval | — | 84.9% |
| LMArena Instruction Following | 1024 | — |
Writing & Preference Granite 4.0 Micro leads
Codellama 70b Instruct: 33.4 (#277), Granite 4.0 Micro: 46.7 (#216)
| Benchmark | Codellama 70b Instruct | Granite 4.0 Micro |
|---|---|---|
| LMArena Text | 1057 | — |
| WildBench | — | 67% |
Frequently asked questions
Is Codellama 70b Instruct better than Granite 4.0 Micro?
Codellama 70b Instruct is the stronger model overall, scoring 33.7 to 29.0 on the Noometry Index.
How many benchmarks do Codellama 70b Instruct and Granite 4.0 Micro share?
0 benchmarks have published results for both models. Codellama 70b Instruct has 7 scored results on Noometry and Granite 4.0 Micro has 8.