Model comparison

Granite 4.1 8b vs Pixtral Large

Granite 4.1 8b is the stronger model overall, scoring 37.4 to 32.2 on the Noometry Index.

Last verified . 0 shared benchmarks.

Granite 4.1 8b IBM

37.4

Rank #205 Confirmed

Pixtral Large Mistral AI

32.2

Rank #259 Reported

Summary

  • The widest gap is in writing & preference, where Granite 4.1 8b leads 48.1 to 32.9.

Side by side

Granite 4.1 8b and Pixtral Large specifications
Granite 4.1 8bPixtral Large
ProviderIBMMistral AI
Noometry Index37.432.2
Released—2024-11-01
WeightsOpenOpen
Context window—128K
Max output—128K
Input $ / M tokens—$2
Output $ / M tokens—$6
Results tracked133

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Granite 4.1 8b: 30.1 (#297), Pixtral Large: —

Coding benchmarks
BenchmarkGranite 4.1 8bPixtral Large
LMArena WebDev1192—
LMArena Coding1312—

Reasoning Granite 4.1 8b leads

Granite 4.1 8b: 25.7 (#143), Pixtral Large: 21.7 (#218)

Reasoning benchmarks
BenchmarkGranite 4.1 8bPixtral Large
EnigmaEval—0.8%
LMArena Hard Prompts1293—

Math Not comparable

Granite 4.1 8b: 36.4 (#166), Pixtral Large: —

Math benchmarks
BenchmarkGranite 4.1 8bPixtral Large
LMArena Math1312—

Knowledge Not comparable

Granite 4.1 8b: 36.1 (#174), Pixtral Large: —

Knowledge benchmarks
BenchmarkGranite 4.1 8bPixtral Large
LMArena Expert1309—

Multimodal Not comparable

Granite 4.1 8b: —, Pixtral Large: 30.6 (#111)

Multimodal benchmarks
BenchmarkGranite 4.1 8bPixtral Large
LMArena Vision—1089

Multilingual Not comparable

Granite 4.1 8b: 41.7 (#204), Pixtral Large: —

Multilingual benchmarks
BenchmarkGranite 4.1 8bPixtral Large
LMArena Non-English1261—
LMArena Chinese1337—
LMArena Russian1240—

Instruction Following Not comparable

Granite 4.1 8b: 66.9 (#203), Pixtral Large: —

Instruction Following benchmarks
BenchmarkGranite 4.1 8bPixtral Large
LMArena Instruction Following1269—

Long Context Not comparable

Granite 4.1 8b: 38.7 (#193), Pixtral Large: —

Long Context benchmarks
BenchmarkGranite 4.1 8bPixtral Large
LMArena Longer Query1275—

Writing & Preference Granite 4.1 8b leads

Granite 4.1 8b: 48.1 (#204), Pixtral Large: 32.9 (#278)

Writing & Preference benchmarks
BenchmarkGranite 4.1 8bPixtral Large
LMArena Text1290—
LMArena Creative Writing1251—
EQ-Bench Creative Writing—988
LMArena Multi-Turn1268—

Frequently asked questions

Is Granite 4.1 8b better than Pixtral Large?

Granite 4.1 8b is the stronger model overall, scoring 37.4 to 32.2 on the Noometry Index.

How many benchmarks do Granite 4.1 8b and Pixtral Large share?

0 benchmarks have published results for both models. Granite 4.1 8b has 13 scored results on Noometry and Pixtral Large has 3.

Related comparisons

Go deeper