Model comparison

Granite 4.2 30b vs Mistral Medium 3.5

Granite 4.2 30b is the stronger model overall, scoring 41.8 to 40.2 on the Noometry Index.

Last verified . 11 shared benchmarks.

Granite 4.2 30b IBM

41.8

Rank #130 Confirmed

Mistral Medium 3.5 Mistral AI

40.2

Rank #152 Confirmed

Summary

  • They share 11 benchmarks with published results for both. Granite 4.2 30b scores higher in 2 categories and Mistral Medium 3.5 in 5 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Granite 4.2 30b leads 27.8 to 17.3.

Side by side

Granite 4.2 30b and Mistral Medium 3.5 specifications
Granite 4.2 30bMistral Medium 3.5
ProviderIBMMistral AI
Noometry Index41.840.2
Released——
WeightsOpenOpen
Context window—262K
Max output—210K
Input $ / M tokens—$1.50
Output $ / M tokens—$7.50
Results tracked1122

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Granite 4.2 30b leads

Granite 4.2 30b: 41.0 (#126), Mistral Medium 3.5: 36.0 (#213)

Coding benchmarks
BenchmarkGranite 4.2 30bMistral Medium 3.5
LMArena Coding13961461
LMArena WebDev—1264

Reasoning Granite 4.2 30b leads

Granite 4.2 30b: 27.8 (#112), Mistral Medium 3.5: 17.3 (#295)

Reasoning benchmarks
BenchmarkGranite 4.2 30bMistral Medium 3.5
LMArena Hard Prompts13741436
Kagi LLM Benchmark—41.4%
NYT Connections (extended)—12.9%
Epoch Capabilities Index—141.35

Math Not comparable

Granite 4.2 30b: —, Mistral Medium 3.5: 39.1 (#113)

Math benchmarks
BenchmarkGranite 4.2 30bMistral Medium 3.5
LMArena Math—1431

Knowledge Too close to call

Granite 4.2 30b: 39.1 (#138), Mistral Medium 3.5: 40.0 (#126)

Knowledge benchmarks
BenchmarkGranite 4.2 30bMistral Medium 3.5
LMArena Expert14061432

Multimodal Not comparable

Granite 4.2 30b: —, Mistral Medium 3.5: 38.3 (#65)

Multimodal benchmarks
BenchmarkGranite 4.2 30bMistral Medium 3.5
LMArena Vision—1223

Multilingual Mistral Medium 3.5 leads

Granite 4.2 30b: 47.3 (#151), Mistral Medium 3.5: 51.9 (#100)

Multilingual benchmarks
BenchmarkGranite 4.2 30bMistral Medium 3.5
LMArena Non-English13401404
LMArena Chinese14141442
LMArena Russian13431395
LMArena French—1448
LMArena German—1451
LMArena Korean—1385
LMArena Spanish—1409

Instruction Following Mistral Medium 3.5 leads

Granite 4.2 30b: 71.2 (#155), Mistral Medium 3.5: 74.6 (#90)

Instruction Following benchmarks
BenchmarkGranite 4.2 30bMistral Medium 3.5
LMArena Instruction Following13471415

Long Context Mistral Medium 3.5 leads

Granite 4.2 30b: 41.4 (#140), Mistral Medium 3.5: 43.2 (#103)

Long Context benchmarks
BenchmarkGranite 4.2 30bMistral Medium 3.5
LMArena Longer Query13591415

Writing & Preference Mistral Medium 3.5 leads

Granite 4.2 30b: 53.8 (#156), Mistral Medium 3.5: 58.5 (#117)

Writing & Preference benchmarks
BenchmarkGranite 4.2 30bMistral Medium 3.5
LMArena Text13611421
LMArena Creative Writing12881374
LMArena Multi-Turn13391423
EQ-Bench 4—993

Frequently asked questions

Is Granite 4.2 30b better than Mistral Medium 3.5?

Granite 4.2 30b is the stronger model overall, scoring 41.8 to 40.2 on the Noometry Index.

Is Granite 4.2 30b or Mistral Medium 3.5 better for coding?

Granite 4.2 30b scores higher on coding benchmarks: 41.0 versus 36.0 in the Noometry coding category.

How many benchmarks do Granite 4.2 30b and Mistral Medium 3.5 share?

11 benchmarks have published results for both models. Granite 4.2 30b has 11 scored results on Noometry and Mistral Medium 3.5 has 22.

Related comparisons

Go deeper