Model comparison

Granite 3.0 8b Instruct vs Mistral Medium

Mistral Medium is the stronger model overall, scoring 36.3 to 31.6 on the Noometry Index.

Last verified . 12 shared benchmarks.

Granite 3.0 8b Instruct IBM

31.6

Rank #270 Confirmed

Mistral Medium Mistral AI

36.3

Rank #218 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Granite 3.0 8b Instruct scores higher in 2 categories and Mistral Medium in 6 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Mistral Medium leads 60.0 to 31.1.

Side by side

Granite 3.0 8b Instruct and Mistral Medium specifications
Granite 3.0 8b InstructMistral Medium
ProviderIBMMistral AI
Noometry Index31.636.3
Released—2023-12-11
WeightsOpenOpen
Context window—262K
Max output—262K
Input $ / M tokens—$1.50
Output $ / M tokens—$7.50
Results tracked1436

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mistral Medium leads

Granite 3.0 8b Instruct: 29.7 (#301), Mistral Medium: 34.2 (#243)

Coding benchmarks
BenchmarkGranite 3.0 8b InstructMistral Medium
LMArena Coding11121434
FrontierCode—8%
SciCode—40.2%
WeirdML—43.7%
BigCodeBench Instruct29.3%—
BigCodeBench Complete35.4%—
ALE-Bench—763.98

Agentic & Tool Use Not comparable

Granite 3.0 8b Instruct: —, Mistral Medium: 28.3 (#90)

Agentic & Tool Use benchmarks
BenchmarkGranite 3.0 8b InstructMistral Medium
Berkeley Function Calling Leaderboard—37.7%

Reasoning Mistral Medium leads

Granite 3.0 8b Instruct: 20.9 (#230), Mistral Medium: 24.0 (#167)

Reasoning benchmarks
BenchmarkGranite 3.0 8b InstructMistral Medium
LMArena Hard Prompts10921426
Kagi LLM Benchmark—50%
CritPt—0%
DTBench—75.5%
LMCA—26.1%
Surface Evolver Bench—26.9%

Math Granite 3.0 8b Instruct leads

Granite 3.0 8b Instruct: 32.8 (#210), Mistral Medium: 28.1 (#245)

Math benchmarks
BenchmarkGranite 3.0 8b InstructMistral Medium
LMArena Math11431408
OTIS Mock AIME 2024-2025—32.2%
ProofBench—9%
MATH Level 5—81.6%
FrontierMath (Feb 2025 set)—0.3%

Knowledge Granite 3.0 8b Instruct leads

Granite 3.0 8b Instruct: 29.6 (#236), Mistral Medium: 25.0 (#265)

Knowledge benchmarks
BenchmarkGranite 3.0 8b InstructMistral Medium
LMArena Expert10871408
GPQA Diamond—59.5%
Humanity's Last Exam—4.5%
Vectara Hallucination Rate—22.7%

Multimodal Not comparable

Granite 3.0 8b Instruct: —, Mistral Medium: 35.3 (#88)

Multimodal benchmarks
BenchmarkGranite 3.0 8b InstructMistral Medium
LMArena Vision—1172

Multilingual Mistral Medium leads

Granite 3.0 8b Instruct: 27.2 (#276), Mistral Medium: 52.1 (#91)

Multilingual benchmarks
BenchmarkGranite 3.0 8b InstructMistral Medium
LMArena Non-English10371408
LMArena Chinese10631447
LMArena Russian10601411
LMArena French—1459
LMArena German—1432
LMArena Japanese—1378
LMArena Korean—1380
LMArena Spanish—1433

Instruction Following Mistral Medium leads

Granite 3.0 8b Instruct: 56.0 (#276), Mistral Medium: 73.7 (#116)

Instruction Following benchmarks
BenchmarkGranite 3.0 8b InstructMistral Medium
LMArena Instruction Following10881398

Long Context Mistral Medium leads

Granite 3.0 8b Instruct: 34.0 (#252), Mistral Medium: 42.9 (#114)

Long Context benchmarks
BenchmarkGranite 3.0 8b InstructMistral Medium
LMArena Longer Query11221406

Writing & Preference Mistral Medium leads

Granite 3.0 8b Instruct: 31.1 (#285), Mistral Medium: 60.0 (#103)

Writing & Preference benchmarks
BenchmarkGranite 3.0 8b InstructMistral Medium
LMArena Text10961424
LMArena Creative Writing10711391
LMArena Multi-Turn10631418
Short-Story Creative Writing—77.3%

Frequently asked questions

Is Granite 3.0 8b Instruct better than Mistral Medium?

Mistral Medium is the stronger model overall, scoring 36.3 to 31.6 on the Noometry Index.

Is Granite 3.0 8b Instruct or Mistral Medium better for coding?

Mistral Medium scores higher on coding benchmarks: 34.2 versus 29.7 in the Noometry coding category.

How many benchmarks do Granite 3.0 8b Instruct and Mistral Medium share?

12 benchmarks have published results for both models. Granite 3.0 8b Instruct has 14 scored results on Noometry and Mistral Medium has 36.

Related comparisons

Go deeper