Model comparison

Granite 4.2 3b vs Mistral Small 3.2

Granite 4.2 3b is the stronger model overall, scoring 39.4 to 31.2 on the Noometry Index.

Last verified . 0 shared benchmarks.

Granite 4.2 3b IBM

39.4

Rank #169 Confirmed

Mistral Small 3.2 Mistral AI

31.2

Rank #280 Confirmed

Summary

  • The widest gap is in knowledge, where Granite 4.2 3b leads 36.3 to 26.7.

Side by side

Granite 4.2 3b and Mistral Small 3.2 specifications
Granite 4.2 3bMistral Small 3.2
ProviderIBMMistral AI
Noometry Index39.431.2
Released—2025-06-20
WeightsOpenOpen
Context window—256K
Max output—16K
Input $ / M tokens—$0.0938
Output $ / M tokens—$0.25
Results tracked116

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Granite 4.2 3b: 40.0 (#151), Mistral Small 3.2: —

Coding benchmarks
BenchmarkGranite 4.2 3bMistral Small 3.2
LMArena Coding1361—

Reasoning Granite 4.2 3b leads

Granite 4.2 3b: 26.0 (#138), Mistral Small 3.2: 18.1 (#287)

Reasoning benchmarks
BenchmarkGranite 4.2 3bMistral Small 3.2
Kagi LLM Benchmark—40.4%
Chess Puzzles—1%
LMArena Hard Prompts1306—
Epoch Capabilities Index—131.74

Math Not comparable

Granite 4.2 3b: —, Mistral Small 3.2: 26.3 (#260)

Math benchmarks
BenchmarkGranite 4.2 3bMistral Small 3.2
OTIS Mock AIME 2024-2025—30.3%

Knowledge Granite 4.2 3b leads

Granite 4.2 3b: 36.3 (#171), Mistral Small 3.2: 26.7 (#256)

Knowledge benchmarks
BenchmarkGranite 4.2 3bMistral Small 3.2
GPQA Diamond—49.1%
LMArena Expert1315—

Multilingual Not comparable

Granite 4.2 3b: 42.1 (#198), Mistral Small 3.2: —

Multilingual benchmarks
BenchmarkGranite 4.2 3bMistral Small 3.2
LMArena Non-English1268—
LMArena Chinese1269—
LMArena Russian1249—

Instruction Following Not comparable

Granite 4.2 3b: 67.1 (#200), Mistral Small 3.2: —

Instruction Following benchmarks
BenchmarkGranite 4.2 3bMistral Small 3.2
LMArena Instruction Following1273—

Long Context Not comparable

Granite 4.2 3b: 39.2 (#185), Mistral Small 3.2: —

Long Context benchmarks
BenchmarkGranite 4.2 3bMistral Small 3.2
LMArena Longer Query1291—

Writing & Preference Granite 4.2 3b leads

Granite 4.2 3b: 47.2 (#212), Mistral Small 3.2: 45.0 (#224)

Writing & Preference benchmarks
BenchmarkGranite 4.2 3bMistral Small 3.2
LMArena Text1293—
LMArena Creative Writing1205—
EQ-Bench Creative Writing—1255
LMArena Multi-Turn1290—

Frequently asked questions

Is Granite 4.2 3b better than Mistral Small 3.2?

Granite 4.2 3b is the stronger model overall, scoring 39.4 to 31.2 on the Noometry Index.

How many benchmarks do Granite 4.2 3b and Mistral Small 3.2 share?

0 benchmarks have published results for both models. Granite 4.2 3b has 11 scored results on Noometry and Mistral Small 3.2 has 6.

Related comparisons

Go deeper