Model comparison

Granite 4.2 30b vs Magistral Small

Granite 4.2 30b is the stronger model overall, scoring 41.8 to 30.2 on the Noometry Index.

Last verified . 0 shared benchmarks.

Granite 4.2 30b IBM

41.8

Rank #130 Confirmed

Magistral Small Mistral AI

30.2

Rank #296 Confirmed

Summary

  • The widest gap is in reasoning, where Granite 4.2 30b leads 27.8 to 6.8.

Side by side

Granite 4.2 30b and Magistral Small specifications
Granite 4.2 30bMagistral Small
ProviderIBMMistral AI
Noometry Index41.830.2
Released—2025-06-10
WeightsOpenOpen
Context window—128K
Max output—40K
Input $ / M tokens—$0.50
Output $ / M tokens—$1.50
Results tracked1110

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Granite 4.2 30b leads

Granite 4.2 30b: 41.0 (#126), Magistral Small: 38.4 (#176)

Coding benchmarks
BenchmarkGranite 4.2 30bMagistral Small
SciCode—35.2%
LMArena Coding1396—

Reasoning Granite 4.2 30b leads

Granite 4.2 30b: 27.8 (#112), Magistral Small: 6.8 (#350)

Reasoning benchmarks
BenchmarkGranite 4.2 30bMagistral Small
ARC-AGI-2—0%
Kagi LLM Benchmark—6.3%
ARC-AGI-1—5%
CritPt—0.3%
Chess Puzzles—3%
LMArena Hard Prompts1374—
DTBench—61.3%
Epoch Capabilities Index—133.19

Math Not comparable

Granite 4.2 30b: —, Magistral Small: 26.2 (#261)

Math benchmarks
BenchmarkGranite 4.2 30bMagistral Small
OTIS Mock AIME 2024-2025—30%

Knowledge Granite 4.2 30b leads

Granite 4.2 30b: 39.1 (#138), Magistral Small: 30.9 (#223)

Knowledge benchmarks
BenchmarkGranite 4.2 30bMagistral Small
GPQA Diamond—56.1%
LMArena Expert1406—

Multilingual Not comparable

Granite 4.2 30b: 47.3 (#151), Magistral Small: —

Multilingual benchmarks
BenchmarkGranite 4.2 30bMagistral Small
LMArena Non-English1340—
LMArena Chinese1414—
LMArena Russian1343—

Instruction Following Not comparable

Granite 4.2 30b: 71.2 (#155), Magistral Small: —

Instruction Following benchmarks
BenchmarkGranite 4.2 30bMagistral Small
LMArena Instruction Following1347—

Long Context Not comparable

Granite 4.2 30b: 41.4 (#140), Magistral Small: —

Long Context benchmarks
BenchmarkGranite 4.2 30bMagistral Small
LMArena Longer Query1359—

Writing & Preference Not comparable

Granite 4.2 30b: 53.8 (#156), Magistral Small: —

Writing & Preference benchmarks
BenchmarkGranite 4.2 30bMagistral Small
LMArena Text1361—
LMArena Creative Writing1288—
LMArena Multi-Turn1339—

Frequently asked questions

Is Granite 4.2 30b better than Magistral Small?

Granite 4.2 30b is the stronger model overall, scoring 41.8 to 30.2 on the Noometry Index.

Is Granite 4.2 30b or Magistral Small better for coding?

Granite 4.2 30b scores higher on coding benchmarks: 41.0 versus 38.4 in the Noometry coding category.

How many benchmarks do Granite 4.2 30b and Magistral Small share?

0 benchmarks have published results for both models. Granite 4.2 30b has 11 scored results on Noometry and Magistral Small has 10.

Related comparisons

Go deeper