Model comparison

Granite 3.1 8b Instruct vs Ministral 8B

Granite 3.1 8b Instruct is the stronger model overall, scoring 32.4 to 28.2 on the Noometry Index.

Last verified . 13 shared benchmarks.

Granite 3.1 8b Instruct IBM

32.4

Rank #258 Confirmed

Ministral 8B Mistral AI

28.2

Rank #325 Confirmed

Summary

  • They share 13 benchmarks with published results for both. Granite 3.1 8b Instruct scores higher in 4 categories and Ministral 8B in 5 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Granite 3.1 8b Instruct leads 31.1 to 12.6.
  • The biggest single-benchmark swing is Berkeley Function Calling Leaderboard: 27.1% for Granite 3.1 8b Instruct and 11.1% for Ministral 8B.

Side by side

Granite 3.1 8b Instruct and Ministral 8B specifications
Granite 3.1 8b InstructMinistral 8B
ProviderIBMMistral AI
Noometry Index32.428.2
Released—2024-10-01
WeightsOpenOpen
Context window—262K
Max output—262K
Input $ / M tokens—$0.15
Output $ / M tokens—$0.15
Results tracked1317

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Granite 3.1 8b Instruct: 34.5 (#233), Ministral 8B: 35.0 (#230)

Coding benchmarks
BenchmarkGranite 3.1 8b InstructMinistral 8B
LMArena Coding11861202

Agentic & Tool Use Granite 3.1 8b Instruct leads

Granite 3.1 8b Instruct: 24.1 (#120), Ministral 8B: 16.4 (#148)

Agentic & Tool Use benchmarks
BenchmarkGranite 3.1 8b InstructMinistral 8B
Berkeley Function Calling Leaderboard27.1%11.1%

Reasoning Granite 3.1 8b Instruct leads

Granite 3.1 8b Instruct: 22.1 (#207), Ministral 8B: 18.4 (#281)

Reasoning benchmarks
BenchmarkGranite 3.1 8b InstructMinistral 8B
LMArena Hard Prompts11451191
DTBench—45.7%

Math Granite 3.1 8b Instruct leads

Granite 3.1 8b Instruct: 33.0 (#209), Ministral 8B: 25.7 (#267)

Math benchmarks
BenchmarkGranite 3.1 8b InstructMinistral 8B
LMArena Math11521188
MATH Level 5—14.9%

Knowledge Granite 3.1 8b Instruct leads

Granite 3.1 8b Instruct: 31.1 (#220), Ministral 8B: 12.6 (#297)

Knowledge benchmarks
BenchmarkGranite 3.1 8b InstructMinistral 8B
LMArena Expert11421170
GPQA Diamond—27.1%
Vectara Hallucination Rate—7.4%

Multilingual Ministral 8B leads

Granite 3.1 8b Instruct: 30.9 (#260), Ministral 8B: 35.1 (#247)

Multilingual benchmarks
BenchmarkGranite 3.1 8b InstructMinistral 8B
LMArena Non-English10991165
LMArena Chinese11451193
LMArena Russian10921195

Instruction Following Ministral 8B leads

Granite 3.1 8b Instruct: 58.6 (#259), Ministral 8B: 60.5 (#250)

Instruction Following benchmarks
BenchmarkGranite 3.1 8b InstructMinistral 8B
LMArena Instruction Following11311161

Long Context Ministral 8B leads

Granite 3.1 8b Instruct: 35.2 (#241), Ministral 8B: 36.7 (#227)

Long Context benchmarks
BenchmarkGranite 3.1 8b InstructMinistral 8B
LMArena Longer Query11621212

Writing & Preference Ministral 8B leads

Granite 3.1 8b Instruct: 35.5 (#266), Ministral 8B: 39.6 (#246)

Writing & Preference benchmarks
BenchmarkGranite 3.1 8b InstructMinistral 8B
LMArena Text11501191
LMArena Creative Writing11291175
LMArena Multi-Turn11081166

Frequently asked questions

Is Granite 3.1 8b Instruct better than Ministral 8B?

Granite 3.1 8b Instruct is the stronger model overall, scoring 32.4 to 28.2 on the Noometry Index.

Is Granite 3.1 8b Instruct or Ministral 8B better for coding?

They score almost the same on coding (34.5 vs 35.0); test both on your own repository before choosing.

How many benchmarks do Granite 3.1 8b Instruct and Ministral 8B share?

13 benchmarks have published results for both models. Granite 3.1 8b Instruct has 13 scored results on Noometry and Ministral 8B has 17.

Related comparisons

Go deeper