Model comparison

Granite 4.1 8b vs Mistral Medium 3.1

Granite 4.1 8b is the stronger model overall, scoring 37.4 to 31.9 on the Noometry Index.

Last verified . 0 shared benchmarks.

Granite 4.1 8b IBM

37.4

Rank #205 Confirmed

Mistral Medium 3.1 Mistral AI

31.9

Rank #266 Reported

Summary

  • The widest gap is in reasoning, where Granite 4.1 8b leads 25.7 to 10.6.
  • Granite 4.1 8b has downloadable open weights; the other is API-only.

Side by side

Granite 4.1 8b and Mistral Medium 3.1 specifications
Granite 4.1 8bMistral Medium 3.1
ProviderIBMMistral AI
Noometry Index37.431.9
Released——
WeightsOpenProprietary
Context window—131K
Max output—105K
Input $ / M tokens—$0.40
Output $ / M tokens—$2
Results tracked133

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Granite 4.1 8b: 30.1 (#297), Mistral Medium 3.1: —

Coding benchmarks
BenchmarkGranite 4.1 8bMistral Medium 3.1
LMArena WebDev1192—
LMArena Coding1312—

Reasoning Granite 4.1 8b leads

Granite 4.1 8b: 25.7 (#143), Mistral Medium 3.1: 10.6 (#341)

Reasoning benchmarks
BenchmarkGranite 4.1 8bMistral Medium 3.1
NYT Connections (extended)—6.5%
Thematic Generalization—20.3%
LMArena Hard Prompts1293—

Math Not comparable

Granite 4.1 8b: 36.4 (#166), Mistral Medium 3.1: —

Math benchmarks
BenchmarkGranite 4.1 8bMistral Medium 3.1
LMArena Math1312—

Knowledge Not comparable

Granite 4.1 8b: 36.1 (#174), Mistral Medium 3.1: —

Knowledge benchmarks
BenchmarkGranite 4.1 8bMistral Medium 3.1
LMArena Expert1309—

Multilingual Not comparable

Granite 4.1 8b: 41.7 (#204), Mistral Medium 3.1: —

Multilingual benchmarks
BenchmarkGranite 4.1 8bMistral Medium 3.1
LMArena Non-English1261—
LMArena Chinese1337—
LMArena Russian1240—

Instruction Following Not comparable

Granite 4.1 8b: 66.9 (#203), Mistral Medium 3.1: —

Instruction Following benchmarks
BenchmarkGranite 4.1 8bMistral Medium 3.1
LMArena Instruction Following1269—

Long Context Not comparable

Granite 4.1 8b: 38.7 (#193), Mistral Medium 3.1: —

Long Context benchmarks
BenchmarkGranite 4.1 8bMistral Medium 3.1
LMArena Longer Query1275—

Writing & Preference Mistral Medium 3.1 leads

Granite 4.1 8b: 48.1 (#204), Mistral Medium 3.1: 55.5 (#145)

Writing & Preference benchmarks
BenchmarkGranite 4.1 8bMistral Medium 3.1
LMArena Text1290—
LMArena Creative Writing1251—
EQ-Bench Creative Writing—1476
LMArena Multi-Turn1268—

Frequently asked questions

Is Granite 4.1 8b better than Mistral Medium 3.1?

Granite 4.1 8b is the stronger model overall, scoring 37.4 to 31.9 on the Noometry Index.

How many benchmarks do Granite 4.1 8b and Mistral Medium 3.1 share?

0 benchmarks have published results for both models. Granite 4.1 8b has 13 scored results on Noometry and Mistral Medium 3.1 has 3.

Related comparisons

Go deeper